Site Reliability Specialist
Core
Build resilient solutions and ensure stable, efficient services for game development teams across Ubisoft by improving platform availability, reliability, and performance.
Role type
Site Reliability Engineer (SRE)
Builds
Resilient cloud-based environments, automated deployment workflows, and observability practices for game development platforms.
Domain
Gaming / Cloud Infrastructure / DevOps
Deliverable
production ML models | infrastructure
Required skills
Infrastructure engineering, automation, DevOps practices, GitLab CI/CD, Python, Bash, Go, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana
Preferred skills
AI-assisted engineering tools (GitHub Copilot, Claude Code)
Technologies
GitLab, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana, Python, Bash, Go
Responsibilities
Define and maintain Service Level Objectives (SLOs) and Service Level Indicators (SLIs), design and implement automation solutions, support observability practices including monitoring and incident management, contribute to root cause analysis, optimize deployment workflows, and maintain cloud-based environments.
Seniority
Mid-level, hands-on IC