Spécialiste en fiabilité des sites
Core
Ensuring availability, reliability, and performance of essential platforms and services for game development at Ubisoft.
Role type
Site Reliability Engineer (SRE)
Builds
Resilient solutions for game development platforms and services
Domain
Video game development / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Infrastructure engineering, DevOps practices, Python, Bash, Go, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana
Preferred skills
AI-assisted engineering tools (GitHub Copilot, Claude Code)
Technologies
GitLab, GitLab CI/CD, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana
Responsibilities
Define and maintain SLOs and SLIs with service teams; Design and implement automation solutions to improve operational efficiency and service reliability; Document technical solutions and support their integration; Collaborate with development, infrastructure, and platform teams to improve operational consistency; Support observability practices including monitoring, logging, alerting, and incident management; Participate in root cause analysis and continuous improvement initiatives after service incidents; Optimize deployment processes and operations through automation and infrastructure improvements; Support the evolution and maintenance of cloud environments and services.
Seniority
Mid-level, hands-on IC