Site Reliability Engineering Manager
Core
Lead a distributed Site Reliability Engineering team to pioneer and prove new approaches to large-scale infrastructure, operations, and managed application services for over 60 million Ubuntu users.
Role type
Senior IC Site Reliability Engineering Manager (DevOps/GitOps)
Builds
Next-generation model-driven operations, infrastructure-as-code automation, and managed application services for cloud-native environments.
Domain
Open-source software, Linux, Cloud-native, Kubernetes, IoT, AI
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Linux operations expertise, DevOps team management, Infrastructure as Code, Cloud topology design, Distributed systems architecture, Agile practices, Service Level Agreement management, Code quality assurance, Testing methodologies, Team mentorship
Preferred skills
Ubuntu system administration, Managing distributed teams
Technologies
Kubernetes, Linux, Cloud platforms
Responsibilities
Lead team in daily agile devops practices, Represent the IS team to stakeholders and customers, Organize and drive internal projects, Mentor engineers to improve skills, Identify and measure team health indicators, Implement structured engineering and operations processes, Ensure team focus on priorities and milestones, Meet service level agreements with global customer deployments, Deliver quality managed services consistently
Seniority
Senior, hands-on IC with management responsibilities