Senior Site Reliability Engineer
Core
Design, build, and operate highly reliable infrastructure and AI-powered workflows for a global AI contacts platform serving millions of users.
Role type
Senior Site Reliability Engineer (AI-native infrastructure)
Builds
Resilient cloud infrastructure, CI/CD pipelines, observability systems, and AI-driven operational automation
Domain
AI / SaaS / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Distributed systems, cloud infrastructure, high-availability architecture, monitoring and observability, incident management, CI/CD pipelines, containerized workloads, AI tool fluency
Preferred skills
Mentoring engineers, cross-functional leadership, proactive system improvement
Technologies
GCP, Claude Code, Cursor, GitHub Copilot
Responsibilities
Define and own SLOs/SLAs/error budgets, evolve observability and incident response practices, improve CI/CD pipelines, identify reliability bottlenecks, build internal AI-powered tooling, champion engineering best practices
Seniority
Senior, hands-on IC with mentorship