Software Engineer-II SRE
Core
Build, operate, and continuously improve reliable production systems and infrastructure at scale.
Role type
Site Reliability Engineer (SRE)
Builds
Production infrastructure, CI/CD pipelines, and automated operational workflows on AWS.
Domain
Cloud Infrastructure & Site Reliability Engineering
Required skills
AWS, Kubernetes, Terraform, CI/CD, GitOps, Python, Go, Linux, Networking, Observability
Preferred skills
EKS, GitHub Actions, Jenkins, GitLab CI, ArgoCD, Datadog, Prometheus, Grafana, CloudWatch
Technologies
AWS, Kubernetes, EKS, Terraform, Python, Go, Bash, GitHub Actions, Jenkins, GitLab CI, ArgoCD, Datadog, Prometheus, Grafana, CloudWatch
Responsibilities
Build and operate production infrastructure on AWS; Manage and automate infrastructure using Terraform; Build and improve CI/CD and GitOps workflows; Monitor and troubleshoot production systems; Participate in on-call, incident response, RCA, and long-term remediation; Contribute to SLIs/SLOs and reliability improvements; Identify opportunities to improve scalability, reliability, and operational efficiency.
Seniority
Mid-level, hands-on IC