Site Reliability Engineer II
Core
Ensure production code quality, ownership of production, and deployments for cloud gaming services and network infrastructure.
Role type
Senior Site Reliability Engineer (Cloud Gaming & Network Services)
Builds
Cloud gaming platforms, network services, and API gateways for PlayStation ecosystem
Domain
Gaming, Cloud Computing, Network Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Linux systems administration, API gateway management (Kong), Service Mesh (Istio, Linkerd, Kuma), traffic management, container orchestration (Kubernetes), distributed data storage, NoSQL databases, observability tools, release engineering, performance analysis
Preferred skills
Python, Bash, Go, Java, C++, Rust, QA/SDET experience
Technologies
Kong, Istio, Linkerd, Kuma, Kubernetes, Rancher, Ceph, Rook, MongoDB, Redis, Cassandra, ElasticSearch, Kafka, PostgreSQL, MySQL, Prometheus, Grafana
Responsibilities
Operate and troubleshoot API gateways in production, manage service mesh lifecycle and reliability practices, handle incident response and on-call rotations, perform load testing and performance analysis, manage distributed data storage and databases at scale
Seniority
Senior, hands-on IC
