Senior Site Reliability Engineer
Core
Own critical platform services end-to-end and partner across engineering to raise the reliability bar for the entire platform.
Role type
Senior Site Reliability Engineer (Foundations team)
Builds
Foundational platform services, observability stack, and resilient deployment pipelines
Domain
Legal technology / AI-native workspace / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Production system operations, debugging under pressure, software development for automation, systems thinking, observability, incident management, on-call experience, cloud infrastructure, Kubernetes
Responsibilities
Build, ship, and operate foundational platform services with full ownership; Build and maintain a high-signal observability stack (metrics, logs, traces) and translate signals into action; Define and evolve SLIs/SLOs, alerting, and reliability reporting for critical systems; Improve on-call and incident response, including escalation paths, coordination, and post-incident follow-ups; Reduce toil through automation, better tooling, and improved system ergonomics; Partner with product and platform engineers to design resilient systems and improve deployment safety