Senior Site Reliability Engineer
Core
Building and improving the infrastructure platform, self-service components, and observability tools for a global mobility app serving millions of users.
Role type
Senior Site Reliability Engineer
Builds
Self-service infrastructure components, observability platforms, and highly available platform services for internal engineering teams and end-users.
Domain
Mobility / Transportation / Cloud Infrastructure
Deliverable
infrastructure
Required skills
Unix/Linux, networking stack, containers and schedulers, monitoring, logging, CAP theorem, programming in at least one language, automation, observability principles, SLI/SLO/SLA definition, Kubernetes, Terraform, SQL and NoSQL databases, GitLab CI/CD.
Preferred skills
Service mesh implementation, network policies, mentoring other teams on reliability practices, simplifying complex setups.
Technologies
Kubernetes (EKS), Terraform, Grafana, Cortex, SQL, NoSQL, GitLab
Responsibilities
Evolve infrastructure platform and build self-service components; Design and implement tooling for availability, scalability, and observability; Define SLIs, SLOs, and SLAs; Share on-call schedule for owned platform services; Solve platform incidents and build automations to prevent recurrence; Participate in recruiting.
Seniority
Senior, hands-on IC