Site Reliability Engineer 3
Core
Design and build global cloud infrastructure for MongoDB Atlas, managing hundreds of thousands of clusters and processing billions of metrics daily.
Role type
Senior Site Reliability Engineer
Builds
Global distributed database platform (MongoDB Atlas) across AWS, Google Cloud, and Microsoft Azure
Domain
Cloud Infrastructure / Distributed Systems
Deliverable
infrastructure
Required skills
Linux system administration, modern programming languages, web and network protocols, automation tool development, cloud provider expertise
Preferred skills
Kubernetes/container orchestration, observability of distributed systems, CI/CD infrastructure, networking/security/hardware tuning
Technologies
AWS, Google Cloud, Microsoft Azure, Kubernetes
Responsibilities
Design and build infrastructure for a global cloud service; implement and troubleshoot automation and monitoring across multiple cloud providers; optimize infrastructure performance from application to firmware levels; participate in on-call rotation for system resilience; improve infrastructure capabilities for cost, simplicity, and maintainability
Seniority
Senior, hands-on IC