Director, Site Reliability Engineering & Service Enablement
Core
Lead the SRE operating model and reliability transformation for ServiceNow's global infrastructure, focusing on cloud-agnostic modernization, automation, and AI-enabled operations.
Role type
Director, Site Reliability Engineering & Service Enablement
Builds
Cloud-agnostic, cloud-ready production platform for ServiceNow customers
Domain
Enterprise SaaS / Cloud Infrastructure / Distributed Systems
Deliverable
production ML models | infrastructure | product features
Required skills
SRE strategy & operating model design, global engineering leadership, SLI/SLO & error budget governance, service registry & ownership models, multi-cloud modernization (AWS/Azure/GCP), Kubernetes & distributed systems, infrastructure automation (IaC), AI integration in operations, production readiness practices, cross-functional executive influence
Preferred skills
AI-assisted operations & autonomous remediation, agentic technologies, Backstage/CMDB experience, disaster recovery & cloud migration expertise
Technologies
AWS, Azure, GCP, Kubernetes, Backstage, CMDB, Infrastructure as Code, AI/ML tools
Responsibilities
Define and execute SRE strategy across reliability, observability, and automation; lead a global organization of managers and technical leaders; establish enterprise reliability standards and service enablement maturity; drive cloud-agnostic patterns and production-readiness practices; lead incident response and blameless postmortems; influence architecture to simplify operating models
Seniority
Director, strategic leadership & global organization building