Senior Site Reliability Engineer
Core
Design, build, and operate highly reliable infrastructure and AI-powered workflows for a global AI contacts platform serving millions of users.
Role type
Senior Site Reliability Engineer (AI-infrastructure focus)
Builds
Resilient global infrastructure, AI-driven operational automation, and developer tooling
Domain
AI-native consumer platform / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Distributed systems, cloud infrastructure, high-availability architecture, observability, incident management, CI/CD pipelines, containerization, AI tooling fluency
Preferred skills
GCP expertise, mentoring engineers, cross-functional leadership
Technologies
GCP, containerized workloads, AI tools (Claude Code, Cursor, GitHub Copilot)
Responsibilities
Define and own SLOs/SLAs/error budgets, evolve observability and incident response practices, improve CI/CD pipelines, identify reliability bottlenecks, build internal AI-powered tooling, champion engineering best practices
Seniority
Senior, hands-on IC with mentorship responsibilities