Engineering Manager - Observability & Reliability Engineering Obsession (x/f/m)
Core
Lead the Reliability & Observability team to drive the evolution of Doctolib's observability platform, ensuring scalability and reliability for healthcare services.
Role type
Engineering Manager (Site Reliability Engineering)
Builds
Observability platform, critical transversal services (secrets management, IaC), and a world-class SRE team.
Domain
Healthcare technology, Cloud-native infrastructure, Observability
Deliverable
production ML models | product features | infrastructure
Required skills
People leadership, Cloud-native architecture, Observability strategy, Infrastructure as Code, Secrets management, Incident response, Team scaling
Preferred skills
High-scale telemetry pipeline design, Backend programming (Go/Python/Ruby), Cultural transformation
Technologies
AWS, GCP, Kubernetes, Fluent Bit, OpenTelemetry, Loki, Elasticsearch, Prometheus, Thanos, Datadog, Terraform, OpenTofu, HashiCorp Vault, Terraform Enterprise
Responsibilities
Lead, coach, and grow a team of Site Reliability Engineers; Define and evolve the observability strategy; Own the strategy for critical transversal services; Manage on-call experience and incident response; Collaborate with cross-functional teams to align observability with product needs.
Seniority
Manager, hands-on technical leadership