Senior DevSecOps/Site Reliability Engineer (AWS)
Core
Own and maintain highly available production systems, lead incident response, and drive operational excellence for scalable cloud infrastructure.
Role type
Senior IC DevSecOps/Site Reliability Engineer (AWS)
Builds
Scalable cloud infrastructure on AWS, Kubernetes (EKS) platforms, and automated CI/CD pipelines
Domain
Cloud Infrastructure / Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
AWS, Terraform, Kubernetes (EKS), GitLab CI, ArgoCD, Linux, cloud networking, Python/Bash/Go scripting, observability (ELK, Prometheus)
Preferred skills
AWS SysOps/DevOps certification, CKA/CKAD, ITSM/change management experience in regulated industries
Technologies
AWS, Terraform, Kubernetes, EKS, GitLab CI, ArgoCD, PagerDuty, Zenduty, Prometheus, Grafana, ELK, Datadog
Responsibilities
Lead incident response (P1/P2) and conduct RCA/PIRs; Design and manage scalable cloud infrastructure; Develop and optimize CI/CD and GitOps pipelines; Manage observability and on-call operations; Collaborate on cloud architecture and compliance initiatives; Mentor team members
Seniority
Senior, hands-on IC