Senior SRE - Platform (Managed Kubernetes Infrastructure)
Core
Designing, building, scaling, and maturing a multi-cloud Kubernetes platform to host internal and external services like Elastic Cloud and Serverless, ensuring global infrastructure reliability.
Role type
Senior Site Reliability Engineer (Platform Engineering)
Builds
Multi-cloud Kubernetes infrastructure, automation tooling, and software supporting product deployment
Domain
Cloud Infrastructure / Kubernetes / SaaS
Deliverable
infrastructure
Required skills
Kubernetes, Golang, Public Cloud Service Providers, Linux, Infrastructure-as-Code, Alerting and Incident Management
Preferred skills
Crossplane, Terraform, Docker, Elastic Stack, Coaching and Mentoring
Technologies
Kubernetes, Golang, Crossplane, Terraform, Docker, Elastic Stack, Prometheus, Influx
Responsibilities
Lead technical initiatives to automate system engineering for global infrastructure reliability; Develop and maintain software, tooling, and automations to scale platform infrastructure; Respond to and prevent repeated customer impact during major incidents; Collaborate in an inclusive environment focusing on operational excellence.
Seniority
Senior, hands-on IC