CareerPlanGet AI match score →

Spécialiste en fiabilité des sites

Montreal, QC, ca💼 Full-time🗓 2026-07-10 → 2026-08-01

Core

Ensuring availability, reliability, and performance of essential platforms and services for game development at Ubisoft.

Role type

Site Reliability Engineer (SRE)

Builds

Resilient solutions for game development platforms and services

Domain

Video game development / Cloud infrastructure

Deliverable

production ML models | infrastructure

Required skills

Infrastructure engineering, DevOps practices, Python, Bash, Go, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana

Preferred skills

AI-assisted engineering tools (GitHub Copilot, Claude Code)

Technologies

GitLab, GitLab CI/CD, Terraform, Kubernetes, AWS, Azure, Ansible, Chef, Prometheus, Grafana

Responsibilities

Define and maintain SLOs and SLIs with service teams; Design and implement automation solutions to improve operational efficiency and service reliability; Document technical solutions and support their integration; Collaborate with development, infrastructure, and platform teams to improve operational consistency; Support observability practices including monitoring, logging, alerting, and incident management; Participate in root cause analysis and continuous improvement initiatives after service incidents; Optimize deployment processes and operations through automation and infrastructure improvements; Support the evolution and maintenance of cloud environments and services.

Seniority

Mid-level, hands-on IC

Sourced via smartrecruiters · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on SmartRecruiters ↗