IN_Associate_SRE _GCC_Advisory_Hyderabad
Core
Site Reliability Engineer ensuring system reliability, availability, performance, and operational excellence for highly scalable, global platforms.
Role type
Associate Site Reliability Engineer (SRE)
Builds
Global-scale production platforms and cloud infrastructure
Domain
Cloud Infrastructure & Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Linux/Unix systems, Python/Go/Java/Bash, Cloud platforms (Azure/AWS/GCP), Kubernetes, Docker, CI/CD pipelines, Monitoring & observability tools, Infrastructure as Code
Preferred skills
Global Capability Center experience, 24x7 production support, SRE best practices, Cloud/Kubernetes certifications, FinOps
Technologies
Azure, AWS, GCP, Kubernetes, Docker, Terraform, ARM, CloudFormation, Pulumi, Prometheus, Grafana, Datadog, ELK, Azure Monitor
Responsibilities
Ensure high availability and reliability of large-scale systems, Monitor and resolve production incidents, Define and track SLIs/SLOs/error budgets, Perform root cause analysis, Build and maintain automation scripts, Support and optimize cloud environments, Improve system resilience through capacity planning
Seniority
Associate, hands-on IC