Site Reliability Engineer (Senior or Staff), Atlas
Core
Design, build, and maintain a reliable, resilient multi-cloud platform hosting business-critical applications for a wide range of customers.
Role type
Senior Site Reliability Engineer (SRE)
Builds
A globally distributed, multi-cloud database platform (MongoDB Atlas) across AWS, Google Cloud, and Microsoft Azure.
Domain
Cloud-native database infrastructure, multi-cloud environments, Linux systems.
Deliverable
production ML models | product features | infrastructure
Required skills
Linux fundamentals, modern programming languages (Go, Ruby, Python), web and network protocols (HTTP, TLS, DNS), multi-cloud operations (AWS, Azure, GCP), automation, system design at scale.
Preferred skills
None stated.
Technologies
AWS, Azure, GCP, Linux, Go, Ruby, Python, HTTP, TLS, DNS.
Responsibilities
Design and build complex systems for the Atlas platform, collaborate with service-owning teams to solve technical challenges and build tooling, participate in 24/7 on-call rotation to resolve disruptions.
Seniority
Senior, hands-on IC