Engineering Manager, Reliability Platform
Core
Design, build, and operate services and infrastructure for DoorDash's Reliability Platform, enabling users and agents to reason about service health, facilitate change control safety, and rapidly address unexpected states.
Role type
Engineering Manager, Reliability Platform
Builds
SLO frameworks, analytics tools, AI Agent enablement, self-service provisioning orchestration, incident management tools, and runtime configuration key-value tooling.
Domain
Food delivery / Cloud Infrastructure / Reliability Engineering
Deliverable
production ML models | infrastructure
Required skills
Leading high-caliber engineering teams, designing complex backend systems, cloud infrastructure fundamentals (AWS, containerization, IaC), SRE concepts (SLOs, error budgets), AI tool adoption, influencing organizational processes, budget management.
Preferred skills
Experience with agentic/AI-assisted workflows, platform mindset, automating toil, global team alignment.
Technologies
AWS, Kubernetes, MCP, AI Agents
Responsibilities
Recruit and retain world-class engineering talent, manage team performance and coaching, establish project execution rituals, align with European counterparts on global culture and planning, own global incident response policies, manage team budget for cloud and vendor spend.
Seniority
Senior, hands-on IC leader