Director of Cloud SRE
Core
Lead a team of engineering leaders and engineers to federate SRE principles and drive strategy for internal observability tooling across global hybrid environments.
Role type
Director of Cloud SRE (Engineering Leadership)
Builds
Internally-owned, vendor-agnostic observability and reliability tooling integrated into CI/CD pipelines.
Domain
Cloud-native and on-premise infrastructure, Site Reliability Engineering, Observability
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
SRE principles (SLIs/SLOs, error budgets, incident management), platform engineering, people leadership, hybrid cloud architecture, observability platform design, CI/CD integration, vendor strategy
Preferred skills
GCP expertise, Agentic AI platform experience, manufacturing/OT environment reliability
Technologies
OpenTelemetry, Dynatrace, GCP, CI/CD pipelines
Responsibilities
Define multi-year strategy for unified observability across cloud and on-premise environments; Lead and grow a team of engineering managers and ICs; Drive roadmap for internally-built tooling with vendor-agnostic architecture; Federate SRE practices across application and platform teams; Embed observability into developer workflows; Extend reliability discipline into manufacturing and OT environments; Represent SRE direction to senior leadership; Contribute to vendor and technology decisions; Maintain hands-on technical credibility.
Seniority
Director, hands-on IC with leadership