Staff Engineer, Site Reliability Engineer
Core
Lead the design and implementation of scalable, fault-tolerant, and observable infrastructure supporting OnStar mobile and web experiences, in-vehicle services, and backend platforms.
Role type
Staff Site Reliability Engineer (technical leader)
Builds
Cloud-native infrastructure, CI/CD pipelines, observability dashboards, and reliability practices for OnStar connected services.
Domain
Automotive connected services, cloud operations, SRE
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Kubernetes, CI/CD pipelines, Python/Go/Java, cloud-native system architecture, incident response, root cause analysis, mentoring, strategic planning
Preferred skills
Privacy engineering, data security, compliance frameworks, AWS/GCP/Azure
Technologies
Prometheus, Grafana, Datadog, Kubernetes, AWS, GCP, Azure
Responsibilities
Lead design of scalable infrastructure, champion configuration management and testing frameworks, partner on service reliability and incident response, drive observability strategy, own on-call practices and postmortems, mentor engineers, support compliance and privacy initiatives
Seniority
Staff, hands-on IC with leadership and mentorship