CareerPlanSign in

Spécialiste Fiabilité outils et infrastructures - TGQF

Montreal, QC, ca💼 Full-time🗓 2026-09-25 → 2026-09-26

Core

Ensure the viability, stability, and performance of tools and infrastructures supporting studio activities, acting as a technical resource for incident prevention and resolution.

Role type

Senior IC DevOps/Infrastructure Reliability Engineer

Builds

Stable and reliable development environments, automated deployment pipelines, and incident management processes

Domain

Video game development, Cloud infrastructure, Observability

Deliverable

production ML models | infrastructure

Required skills

Software development, System administration, Database administration, Infrastructure automation (Cloud & On-prem), Programming languages, Observability technologies, CI/CD processes, Cloud services, Network infrastructure, Configuration management, Linux/Windows administration, Redundant and scalable architecture design, Code optimization, Task automation

Preferred skills

Agile methodologies, Technical teaching and mentoring, Proactive problem solving, Rapid adaptation to fast-paced environments

Technologies

Grafana, Splunk, Elasticsearch, Prometheus, OpenTelemetry, Docker, Git, Terraform, Ansible, Chef

Responsibilities

Accompany development teams in technology choices to improve system visibility and robustness; Automate processes to facilitate team work; Implement tools and methods for secure service deployment and incident management; Diagnose and permanently fix anomalies and infrastructure failures; Coordinate resources to restore service level objectives; Design, deploy, secure, and maintain reliable environments; Provide continuous technical support and proactively address issues; Create and maintain deployment guides and technical documentation.

Seniority

Senior, hands-on IC

Sourced via smartrecruiters · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.