Senior Site Reliability Engineer
Core
Own operational excellence and reliability for Cloud Infrastructure, ensuring resilient systems and automating high-risk manual processes.
Role type
Senior Site Reliability Engineer (Cloud Infrastructure)
Builds
AWS, Kubernetes, MongoDB, Temporal, observability stack, and vendor integrations
Domain
Cloud Infrastructure / SaaS Platform
Deliverable
production ML models | infrastructure
Required skills
AWS, Kubernetes, MongoDB, Python scripting, incident management, automation, FinOps, observability
Preferred skills
Temporal, vendor management, cost optimization
Responsibilities
Manage incident response and post-mortems, automate low-frequency high-consequence operations, own a specific platform domain, manage third-party vendor relationships, optimize cloud costs
Seniority
Senior, hands-on IC
Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.