SRE Especialista - Gcp (Site Reliability Engineer)
Core
Design, maintain, and secure Google Cloud Platform infrastructure for a large energy group's distributed systems and services.
Role type
Senior Site Reliability Engineer (GCP)
Builds
Production cloud infrastructure, containerized workloads, and automated deployment pipelines for energy distribution and fintech services.
Domain
Energy sector / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Google Cloud Platform (GCP), Kubernetes (GKE), Terraform, Infrastructure as Code, IAM, Service Accounts, Secret Manager, VPC networking, Docker, Observability, Incident response, Bash/Python/Go scripting, DevSecOps practices
Preferred skills
Cloud SQL, Pub/Sub, Memorystore, Artifact Registry, Datadog, OpenTelemetry, Policy as Code (OPA/Gatekeeper), Kubernetes hardening, SBOM, FinOps, regulated industry experience, Cloud DevOps Engineer/Architect certification
Technologies
GCP, GKE, Cloud Run, Terraform, Docker, Datadog, OpenTelemetry, OPA, Gatekeeper
Responsibilities
Design and maintain GCP infrastructure; Administer GKE, Cloud Run, and managed services; Build and evolve reusable Terraform modules; Define network, IAM, and secret management standards; Implement secure deployment pipelines; Establish observability with metrics, logs, and tracing; Support SLI/SLO definition and capacity planning; Automate provisioning and operational tasks; Investigate incidents and eliminate root causes; Implement security controls for images and workloads; Monitor platform costs and operational risks; Document standards and guide teams on secure infrastructure usage
Seniority
Senior, hands-on IC