Développeur(se) en fiabilité de site
Core
Build and maintain the reliability, observability, and autonomy of a global AI-powered asset management platform serving 14,000+ industrial customers.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud-native asset management platform for industrial operations
Domain
Industrial IoT / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Distributed systems observability, SRE concepts (SLOs, error budgets, incident management), Infrastructure as Code, Cloud-native platforms, Software development
Preferred skills
TypeScript, Node.js
Technologies
Cloud-native platforms, Infrastructure as Code tools
Responsibilities
Evaluate service maturity and provide recommendations to development teams, Implement best practices for observability, Enable developer autonomy in deployment and operations, Mentor developers on reliability practices, Collaborate with platform teams to drive adoption of reliability standards
Seniority
Mid-level, hands-on IC