DevOps / SRE Engineer
Core
Operate mission-critical production systems across hybrid environments, ensuring system reliability, deployment stability, and rapid incident resolution.
Role type
Senior IC DevOps / Site Reliability Engineer
Builds
Production services and integrations
Domain
Cloud infrastructure and hybrid environments
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Kubernetes operations, CI/CD pipeline management, cloud platform administration (AWS/Azure), Infrastructure as Code (Terraform), observability stack management, scripting and automation (Bash, Python, PowerShell)
Preferred skills
API Gateway management, Cloudflare / APIM knowledge, message queue support (RabbitMQ), caching tools (Redis), legacy/Windows-based system support
Technologies
Kubernetes, GitHub Actions, Azure DevOps, Octopus, Terraform, Prometheus, Grafana, AWS, Azure, Bash, Python, PowerShell, RabbitMQ, Redis
Responsibilities
Support 24x7 production systems, participate in on-call rotation, troubleshoot incidents across CI/CD pipelines and Kubernetes clusters, perform incident triage and recovery, ensure safe deployments with rollback mechanisms
Seniority
Senior, hands-on IC