Senior Site Reliability Engineer
Core
Own the infrastructure behind every player connection for 2K's global game services, account platforms, and CI/CD pipelines across AWS, GCP, and on-premises data centers.
Role type
Senior Site Reliability Engineer (Hands-on technical leader)
Builds
Scalable multi-cloud and hybrid infrastructure, Kubernetes platforms, observability stacks, and automated CI/CD pipelines for game services.
Domain
Gaming industry, Cloud Infrastructure, Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Kubernetes (EKS/GKE), Infrastructure as Code (Terraform/Pulumi), Observability (Prometheus/Grafana/Datadog), Linux internals, Go/Python/TypeScript, Incident response, Service mesh (Istio/Cilium)
Preferred skills
Live-service game experience, FinOps, AI/Agentic Development, Mentoring
Technologies
AWS, GCP, Kubernetes, Terraform, Pulumi, ArgoCD, Flux, Istio, Cilium, Prometheus, Grafana, Datadog, GitHub Actions, Jenkins, Ansible, Puppet, OpenTelemetry
Responsibilities
Design and operate scalable multi-cloud infrastructure; Build and run full observability stack; Lead chaos engineering and incident response; Eliminate toil through automation and self-service provisioning; Promote SRE practices across studios.
Seniority
Senior, hands-on IC
