Senior Site Reliability Engineer (Hosted Infra)
Core
Build and operate multi-cloud infrastructure at scale, automating workflows to ensure reliability and efficiency for Elastic Cloud.
Role type
Senior Site Reliability Engineer (Hosted Infra)
Builds
Elastic Cloud infrastructure across 4 cloud providers and 70+ regions
Domain
Cloud Infrastructure / SRE
Deliverable
production ML models | infrastructure
Required skills
Linux systems administration, Golang, containerized workloads, automation, observability, incident management, code review, documentation
Preferred skills
Terraform/OpenTofu, Puppet/OpenVox, Ansible, Argo CD, Argo Workflows, CUE, Docker, Kubernetes, Ubuntu, Elastic Stack, Prometheus, Influx
Technologies
Golang, Linux, Terraform, OpenTofu, Puppet, OpenVox, Ansible, Argo CD, Argo Workflows, CUE, Docker, Kubernetes, Ubuntu, Elastic Stack, Prometheus, Influx
Responsibilities
Engineering software to automate large-scale systems, optimizing host reliability and lifecycle, strengthening observability posture, scaling global infrastructure, participating in on-call rotation and incident response, mentoring teammates
Seniority
Senior, hands-on IC