Site Reliability Engineer (CX)
Core
Design and build custom infrastructure solutions, automate CI/CD pipelines, and maintain production-grade clusters using software engineering principles to enable developer velocity.
Role type
Site Reliability Engineer (Infrastructure Automation)
Builds
Custom automation tools, CI/CD pipelines, and production-grade Kubernetes clusters
Domain
Cloud Infrastructure (Azure/AWS/GCP) and DevOps
Deliverable
production ML models | product features | infrastructure
Required skills
Kubernetes, Infrastructure as Code (Pulumi, Terraform, Ansible), Cloud platforms (Azure, AWS, GCP), Golang, Python, CI/CD pipeline automation, System scalability and resilience design
Preferred skills
Self-hosted Kubernetes experience, Security integration (ISO 27001, SOC2), Observability and incident response processes
Technologies
Pulumi, Terraform, Ansible, Azure, AWS, GCP, Golang, Python
Responsibilities
Design and build custom solutions when existing tools fall short, Debug and improve CI/CD pipelines end-to-end, Work cross-functionally with developers and product teams, Think critically about systems scalability and resilience, Automate infrastructure using reusable code-based solutions
Seniority
Mid-level to Senior, hands-on IC