Senior Site Reliability Engineer
Core
Own reliability, scalability, and security for critical platforms supporting medicines, vaccines, and animal health products worldwide.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Resilient, high-performing systems for global healthcare biopharma technology
Domain
Healthcare biopharma / Cloud infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
SRE/DevOps engineering, Cloud platforms (AWS/Azure/GCP), Infrastructure as Code (Terraform/CloudFormation), CI/CD pipelines, Observability (Prometheus/Grafana/OpenTelemetry), Distributed systems, Networking, Security compliance, Python/Go/Bash scripting, Capacity planning, Incident response, SLO/SLI management, Technical documentation
Preferred skills
Regulated industry experience (pharma/healthcare), Platform engineering patterns, Multi-tenant services, Conversational BI tools, LLM/Copilot concepts, Team leadership
Technologies
AWS, Azure, GCP, Terraform, CloudFormation, Git, Prometheus, Grafana, ELK, OpenTelemetry, Python, Go, Bash, PowerShell, Kubernetes
Responsibilities
Define and own SLOs/SLIs and error budgets, Design scalable architectures and automation, Build and maintain IaC and CI/CD pipelines, Implement observability and monitoring, Lead incident response and postmortems, Establish operational frameworks and runbooks, Mentor SRE and DevOps engineers, Perform capacity planning and troubleshooting, Partner with product and security teams, Produce compliance evidence and technical documentation
Seniority
Senior, hands-on IC with strategic leadership