Staff Engineer, Site Reliability Engineer
Core
Lead the design and implementation of scalable, fault-tolerant, and observable infrastructure supporting OnStar mobile, web, and in-vehicle experiences.
Role type
Staff Site Reliability Engineer (SRE)
Builds
Cloud-native infrastructure, CI/CD pipelines, and observability tooling for millions of customers.
Domain
Automotive connected services / Cloud infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Kubernetes, CI/CD pipelines, Python/Go/Java, Prometheus, Grafana, Datadog, SLO definition, root cause analysis, team mentoring
Preferred skills
Privacy engineering, data security, compliance frameworks
Technologies
AWS, GCP, Azure, Kubernetes, Prometheus, Grafana, Datadog
Responsibilities
Lead design of scalable infrastructure, champion configuration management and testing frameworks, partner on deployment safety and incident response, define observability strategy, manage on-call practices and postmortems, mentor engineers
Seniority
Staff, hands-on IC with strategic leadership