Manager, Cloud Operations Engineering
Core
Lead the Cloud Operations Engineering team to ensure uptime guarantees for the MongoDB Atlas multi-cloud database platform.
Role type
Senior Engineering Manager (Cloud Operations/SRE)
Builds
Internal tools, process automation, and monitoring systems for the Atlas platform
Domain
Cloud infrastructure, distributed databases, SRE
Required skills
Team leadership, incident management, Linux system administration, cloud infrastructure (AWS/GCP/Azure), distributed systems operations, scripting/programming
Preferred skills
Strategic process implementation, performance troubleshooting, cross-functional collaboration
Technologies
Linux, AWS, GCP, Azure, Java, Go, Python, Typescript
Responsibilities
Scale the team through strategic process implementation and tool refinement; Provide career development feedback to direct reports; Monitor team health indicators and performance metrics; Coordinate global on-call rotation for customer incidents; Diagnose live incidents and differentiate platform vs. usage issues; Automate internal processes and routine monitoring tasks; Collaborate with Product Management and Cloud Engineering on infrastructure improvements.
Seniority
Senior, hands-on IC with management responsibilities