Senior Site Reliability Engineer [SRE Infrastructure]
Core
Design, build, and maintain scalable cloud and on-premises infrastructure, ensuring stability and reliability for global data centers and multiple platforms.
Role type
Senior Site Reliability Engineer (Infrastructure)
Builds
Scalable cloud and on-premises infrastructure, observability solutions, deployment platforms, and internal tooling.
Domain
Cloud infrastructure, distributed systems, and gaming technology
Deliverable
production ML models | product features | infrastructure
Required skills
Infrastructure as Code (Terraform, Chef, Ansible), Python/Go/Ruby programming, Kubernetes (EKS, Rancher), distributed systems management, networking fundamentals (DNS, CDNs, VPCs), API gateways, incident response, root cause analysis, mentoring
Preferred skills
None stated
Technologies
Terraform, Chef, Ansible, Python, Go, Ruby, Kubernetes, EKS, Rancher, DNS, CDNs, VPCs
Responsibilities
Design and maintain scalable infrastructure; build observability solutions; develop deployment platforms and tooling; standardize infrastructure patterns; mentor engineers; participate in on-call rotation and incident response; partner with Security and Networking teams.
Seniority
Senior, hands-on IC