CareerPlanSign in

Développeur(se) en fiabilité de site

Montreal💼 Full-time🗓 2026-02-12 → 2026-09-26

Core

Build and maintain the reliability, observability, and autonomy of a global AI-powered asset management platform serving 14,000+ industrial customers.

Role type

Site Reliability Engineer (SRE)

Builds

Cloud-native asset management platform for industrial operations

Domain

Industrial IoT / Cloud Infrastructure

Deliverable

production ML models | infrastructure

Required skills

Distributed systems observability, SRE concepts (SLOs, error budgets, incident management), Infrastructure as Code, Cloud-native platforms, Software development

Preferred skills

TypeScript, Node.js

Technologies

Cloud-native platforms, Infrastructure as Code tools

Responsibilities

Evaluate service maturity and provide recommendations to development teams, Implement best practices for observability, Enable developer autonomy in deployment and operations, Mentor developers on reliability practices, Collaborate with platform teams to drive adoption of reliability standards

Seniority

Mid-level, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.