Senior Site Reliability Engineer (SRE)
Core
Ensuring reliability, scalability, and performance of distributed store technologies powering thousands of pharmacy and retail locations.
Role type
Senior Site Reliability Engineer
Builds
Pharmacy platforms, Point of Sale (POS) systems, handheld devices, store servers, and connectivity infrastructure.
Domain
Retail and Pharmacy Technology
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Observability and monitoring (Splunk, Dynatrace, Datadog, Prometheus, Grafana), SRE best practices, incident response, root cause analysis, cloud-native architectures, Kubernetes, OpenShift, CI/CD automation, microservices, containerization, Java, Python, AWS, Microsoft Azure, Google Cloud, Rancher, Docker, web APIs, source control, GitHub, BitBucket, Jenkins
Preferred skills
Incident Management, Change Management, Infrastructure Support, Problem Management, retail SRE experience, API management (Apigee, Vordel, Data power)
Technologies
Splunk, Dynatrace, Datadog, Prometheus, Grafana, Kubernetes, OpenShift, Rancher, Docker, AWS, Microsoft Azure, Google Cloud, GitHub, BitBucket, Jenkins, Apigee, Vordel, Data power
Responsibilities
Architect and optimize observability solutions; develop proactive monitoring, alerting, and dashboarding strategies; partner with engineering teams to embed SRE best practices; lead incident response, root cause analysis, and reliability reviews; champion cloud-native architectures and microservices; drive deployment excellence through CI/CD automation.
Seniority
Senior, hands-on IC