CareerPlanGet AI match score →

Site Reliability Engineer - I

Bengaluru, Karnataka, India💼 Full-time🗓 2026-07-16 → 2026-07-18

Core

Ensure production services remain available, scalable, and efficient on Google Cloud Platform by bridging development and operations for containerized infrastructure.

Role type

Entry-level Site Reliability Engineer (SRE)

Builds

Containerized applications and infrastructure on Google Cloud Platform

Domain

Cloud infrastructure, networking, and observability

Deliverable

production ML models | infrastructure

Required skills

GitOps (ArgoCD), Linux internals, Kubernetes networking, TCP/IP/HTTP troubleshooting, AIOps concepts, Grafana ecosystem, Python/Bash/Go scripting

Preferred skills

Anomaly detection concepts, log-based ML models

Technologies

ArgoCD, GKE, Grafana, Mimir, Prometheus, Loki, Tempo, Python, Bash, Go

Responsibilities

Deploy and manage containerized application lifecycles via ArgoCD pipelines; Investigate and resolve infrastructure, OS, application, and network alerts; Diagnose connectivity and latency issues across cloud VPCs and Kubernetes overlays; Troubleshoot OS bottlenecks including CPU throttling, memory leaks, and storage constraints; Utilize AI-driven operations tools for event correlation and root-cause analysis; Participate in on-call rotations to mitigate production container issues

Seniority

Entry-level, hands-on IC

Sourced via linkedin · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on LinkedIn ↗