Sr Mgr, Reliability Engineering Mgmt
Core
Lead a global team of SRE leaders, managers, and engineers to ensure the reliability, availability, and cloud modernization of critical enterprise platforms and applications.
Role type
Senior Manager, Site Reliability Engineering
Builds
Global SRE organization, cloud-native architectures, automated operational practices, and resilient production systems
Domain
Enterprise Software / Cloud Infrastructure / Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
People management (managing managers), Cloud modernization strategy, Kubernetes & container orchestration, Observability (SLI/SLO/Error Budgets), CI/CD & Release pipelines, Automation & Infrastructure as Code, Incident & Problem management, Distributed systems architecture
Preferred skills
AI-assisted operations, Legacy platform migration, Cross-functional leadership
Technologies
AWS, GCP, Azure, Kubernetes, Linux, CI/CD tools, Observability platforms
Responsibilities
Lead and develop a global team of SRE leaders, managers, and engineers; Define and drive SRE strategy and operating model; Own and evolve observability capabilities; Lead strategy for production-like staging environments; Partner with Engineering and Release teams to strengthen release pipelines; Drive cloud modernization initiatives; Establish reliability standards and measurable outcomes; Provide leadership during major incidents; Build a global engineering culture.
Seniority
Senior, hands-on IC with people management