PV| Analista SRE Sênior
Core
Build and evolve SRE culture, design and operate highly available cloud environments, and implement automation for provisioning, deployment, and incident response.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Resilient cloud infrastructure and automated operational solutions on AWS
Domain
Cloud Infrastructure / DevOps / FinOps
Deliverable
production ML models | infrastructure
Required skills
AWS infrastructure, Terraform, Infrastructure as Code (IaC), CI/CD pipelines, observability (logs, metrics, tracing), SQL database troubleshooting, FinOps practices
Preferred skills
Kubernetes, container orchestration, Prometheus, Grafana, Datadog, DevSecOps, AIOps, distributed platforms, microservices
Technologies
AWS, Terraform, GitLab, Prometheus, Grafana, Datadog, Kubernetes, SQL
Responsibilities
Design and automate highly available cloud environments, implement and maintain IaC solutions, develop automation for provisioning and incident response, define and monitor SLIs/SLOs/SLAs, manage incident response and root cause analysis, support cloud cost management (FinOps)
Seniority
Senior, hands-on IC