Director, Site Reliability Engineering
Core
Lead a global Site Reliability Engineering organization to ensure the reliability, scalability, and operational excellence of the Aiven cloud data platform.
Role type
Director, Site Reliability Engineering
Builds
Global SRE operating model, incident response frameworks, and automation tooling for a 24/7/365 multi-cloud platform
Domain
Cloud infrastructure, distributed systems, multi-cloud solutions
Deliverable
infrastructure
Required skills
Global SRE organization leadership, reliability strategy definition, incident management, team scaling and mentorship, SLI/SLO design, distributed systems architecture, cross-functional executive partnership, organizational design
Preferred skills
Multi-regional leadership experience, follow-the-sun operational models, data-driven prioritization, large-scale change management
Technologies
Cloud infrastructure, distributed systems, automation frameworks
Responsibilities
Define global SRE operating strategy across regions, build and lead multi-regional SRE teams, set reliability vision and roadmap, own global incident management strategy, establish metrics-driven operating cadence
Seniority
Director, strategic leadership with hands-on technical accountability