Ridehailing, Site Reliability Engineer
Core
Building and running large-scale, fault-tolerant, reliable systems for Waymo's fully autonomous ride-hail service and supporting infrastructure.
Role type
Senior Site Reliability Engineer (Autonomous Vehicle Operations)
Builds
Waymo's fully autonomous ride-hail service, depot logistics, automation flow, and critical vehicle state infrastructure
Domain
Autonomous driving technology / Ride-hailing
Deliverable
production ML models | infrastructure
Required skills
C++, Java, Python, distributed systems architecture, performance profiling, large-scale refactoring, SLI/SLO/SLA framework design, observability systems, incident response, technical leadership, mentoring
Preferred skills
Engineering leadership at scale, data-driven SLO frameworks, complex system failure resolution, Computer Science degree
Technologies
C++, Java, Python
Responsibilities
Manage end-to-end availability and performance for core fleet services; Write designs and implement software to improve system architecture, telemetry, or deployment; Serve as first responder for fleet and supply infrastructure leading incident response; Work with partners to develop technical direction and architectural guidelines; Mentor junior and mid-level engineers
Seniority
Senior, hands-on IC with leadership responsibilities