CareerPlanGet AI match score →

Site Reliability Engineer, Cloud Cost Utilization

United Kingdom🌐 Remote💼 Full-time🗓 2026-05-18 → 2026-07-31

Core

Making cloud spending visible, understandable, and actionable across GitLab's infrastructure by building systems for cost tracking, attribution, and optimization.

Role type

Senior IC Site Reliability Engineer (Cloud Cost Utilization)

Builds

Cloud billing data pipelines, cost anomaly detection workflows, and observability tooling for AWS and GCP.

Domain

Cloud Infrastructure / FinOps

Deliverable

production ML models | product features | dashboards & analysis

Required skills

Cloud cost management (AWS/GCP), FinOps FOCUS specification, resource tagging strategies, Infrastructure as Code (Terraform/Ansible), Observability stacks (Prometheus/LGTM/ELK), Data pipeline development

Preferred skills

Cross-functional collaboration with Finance/Engineering, Explaining technical data to non-engineering audiences, Self-directed work in async environments

Technologies

AWS, GCP, Terraform, Ansible, Prometheus, Loki, Grafana, Tempo, Mimir, ELK, FOCUS

Responsibilities

Design and maintain cloud resource tagging and labeling strategies; Develop tooling to ingest and normalize cloud billing data; Automate cost anomaly detection and forecasting; Contribute to observability stacks to surface cost signals; Partner with leadership on cloud cost forecasting; Act as SME for cloud cost attribution; Collaborate on audits and financial reporting.

Seniority

Senior, hands-on IC

Rewrite
## Responsibilities - Build and improve the systems, standards, and workflows that help teams understand the real cost of the services they run. - Develop resource tagging and labeling approaches, improve billing data quality, and create tooling that supports better decisions across AWS and GCP. - Work through technical and organizational ambiguity, connect infrastructure data with business context, and help teams act on cost signals with confidence. - Build cloud billing data pipelines that normalize multi-cloud cost data using the FinOps Open Cost and Usage Specification (FOCUS). - Improve cloud resource tagging and labeling standards so teams can understand spend by service, environment, and ownership. - Develop cost anomaly detection, forecasting, and alerting workflows that give teams timely insight into infrastructure usage. - Extend observability systems so cost signals can be reviewed alongside reliability and operational data. ## Requirements - Design and maintain cloud resource tagging and labeling strategies across GCP and AWS to support accurate cost attribution. - Develop tooling and pipelines to ingest, normalize, and report on cloud billing data using the FOCUS specification. - Automate cost anomaly detection, forecasting, and alerting so engineering teams can respond quickly to changes in infrastructure spend. - Contribute to GitLab's observability and monitoring stacks, including Prometheus, LGTM (Loki, Grafana, Tempo, and Mimir), and ELK, with a focus on surfacing cost efficiency signals. - Partner with Finance and Engineering leadership to support cloud cost forecasting for planning and budget discussions. - Act as a subject matter expert for cloud cost attribution, tagging strategy, and FOCUS adoption across GitLab Infrastructure. ## Nice to Have - Collaborate with Finance and Compliance teams on audits, c ## Benefits - GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems.
Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗