CareerPlanSign in

SRE (Site Reliability Engineer)

Brasil💼 Full-time🗓 2026-09-04 → 2026-09-21

Core

Maintain production stability and resilience for a high-criticality iGaming platform, responding to incidents with agility and reducing MTTR through automation.

Role type

Senior Site Reliability Engineer (SRE)

Builds

Production environments for iGaming/betting platform

Domain

iGaming / Cloud Infrastructure

Deliverable

production ML models | product features | dashboards & analysis | infrastructure

Required skills

Incident response, Root cause analysis, SLO/Error Budget management, Linux troubleshooting, Distributed systems, CI/CD pipelines, GitOps, Infrastructure as Code, Kubernetes, Cloud providers, Observability tools

Preferred skills

Internal Developer Platform (IDP) experience, Advanced CI/CD tools, Basic backend knowledge, Open Source contributions

Technologies

Terraform, Terragrunt, Ansible, AWS, GCP, Azure, Docker, Kubernetes, Datadog, Grafana, Prometheus, ArgoCD, Flux

Responsibilities

Investigate production incidents and perform root cause analysis, Define and monitor SLIs/SLOs/Error Budgets, Create and evolve CI/CD pipelines and GitOps strategies, Provision and maintain infrastructure using IaC, Manage Docker/Kubernetes environments and implement observability solutions

Seniority

Senior, hands-on IC

Sourced via adzuna · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.