Manager, Site Reliability Engineering
Core
Lead the IDaaS Site Reliability Engineering Group to scale high-throughput, 99.999 availability infrastructure on AWS, managing teams focused on Edge networking, K8s, CI/CD, and observability.
Role type
Manager, Site Reliability Engineering (Technical Leadership)
Builds
Edge networking, K8s platform, CI/CD pipelines, observability platforms, automation tooling for the IDaaS service
Domain
Cloud Infrastructure / Identity & Access Management (IDaaS)
Deliverable
production ML models | product features | infrastructure
Required skills
Technical leadership, people management, Agile/DevOps methodologies, large-scale public cloud infrastructure (AWS), cloud-native architectures, containerization (Kubernetes), Infrastructure as Code (Terraform), CI/CD pipelines, software development, PaaS, automation, observability platforms (Grafana, Splunk, APM)
Preferred skills
Multi-cloud environment experience
Technologies
AWS, Kubernetes, Terraform, Grafana, Splunk, APM
Responsibilities
Manage SRE teams supporting IDaaS workloads; drive microservice journey and DevOps maturity; accelerate engineering velocity via tooling and self-healing patterns; mentor engineers and managers; perform engineering design evaluations; improve SDLC and CI/CD processes; manage service expectations and resource allocation
Seniority
Manager, hands-on technical leadership