Senior Manager, Site Reliability Engineering
Core
Lead the SRE and DevOps function to ensure the reliability, availability, and scalability of an AWS-hosted SaaS platform handling sensitive patient data for biopharmaceutical companies.
Role type
Senior Manager, Site Reliability Engineering
Builds
AWS cloud-native infrastructure, CI/CD pipelines, and observability tooling for a HIPAA/SOC 2 compliant SaaS platform.
Domain
Healthcare technology / Cloud Infrastructure
Deliverable
infrastructure
Required skills
AWS cloud architecture, Infrastructure as Code (AWS CDK, Terraform), SLO/SLI management, incident response leadership, CI/CD pipeline optimization, distributed team management, security compliance (HIPAA, SOC 2, HITRUST)
Preferred skills
Golang, OpenTofu, PySpark
Technologies
AWS (ECS, EC2, Aurora RDS, DynamoDB, Lambda, S3, SQS, EventBridge, Cognito, Secrets Manager, CloudFront), GitHub Actions, CloudWatch, Sentry, Jira
Responsibilities
Define and maintain SLOs/SLIs, lead incident response and on-call operations, architect and maintain IaC infrastructure, manage offshore contractor teams, optimize cloud costs and capacity, partner with security on compliance audits
Seniority
Senior, hands-on IC with people management