Site Reliability Engineer, Infrastructure Platforms — UK (Intermediate to Senior Staff)
Core
Keep GitLab's user-facing services and production systems reliable, scalable, and efficient by building automation, tooling, and infrastructure-as-code workflows.
Role type
Senior IC Site Reliability Engineer (Infrastructure Platforms)
Builds
Production infrastructure, automation tooling, and CI/CD pipelines for GitLab.com and dedicated services
Domain
Cloud Infrastructure / DevOps / SRE
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Kubernetes, Infrastructure as Code (Terraform), Go, Cloud providers (AWS/GCP), Observability (metrics/logs/SLOs), Incident response, Automation engineering
Preferred skills
AI integration for operational efficiency, Kubernetes operators/controllers, Ruby
Technologies
Kubernetes, Terraform, Go, AWS, GCP, GitOps
Responsibilities
Operate and troubleshoot production systems on Kubernetes, Write and maintain infrastructure as code, Participate in on-call and incident response, Build automation to reduce toil, Contribute to the observability stack, Document architecture decisions and runbooks
Seniority
Senior, hands-on IC