Site Reliability Engineer II (SRE)
Core
Guide reliability and performance of infrastructure and software systems, enabling safe and quick deployment of new features for customers.
Role type
Senior IC Site Reliability Engineer
Builds
Reliable and flexible infrastructure solutions (AWS, Kubernetes, CI/CD, observability)
Domain
Cloud Infrastructure & DevOps
Deliverable
production ML models | infrastructure
Required skills
Python, Javascript, Unix Shell, AWS, Terraform, Kubernetes, CI/CD pipelines, Docker
Responsibilities
Develop and maintain cloud infrastructure platforms (AWS accounts, networking, Kubernetes clusters); Develop and maintain observability and monitoring tools (Coralogix, Sentry); Develop and maintain CI/CD tools (GitHub Actions, ArgoCD); Estimate timelines for features and tasks balancing effort, risk, and impact; Take ownership of assigned problem statements and drive them to completion.