Site Reliability Engineer
Core
Deploy, operate, and automate the accesso Horizon product in customer cloud environments (AWS/Azure/GCP) to ensure scalability and reliability.
Role type
Site Reliability Engineer (IC)
Builds
Managed access solutions for attractions and leisure venues
Domain
Cloud Infrastructure & Site Reliability Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Terraform, ArgoCD, CI/CD pipeline management, Kubernetes (EKS/AKS/GKE), Python or Bash scripting, Linux command line, incident triage, root cause analysis, monitoring and alerting configuration
Preferred skills
Cloud platform expertise (AWS/Azure/GCP), self-directed learning, minimal direction autonomy
Technologies
Terraform, ArgoCD, GitHub Actions, Prometheus, Grafana, Coralogix, Dynatrace, Azure SQL, RDS, Kubernetes (AKS, GKE, EKS), Docker
Responsibilities
Provision and deploy Horizon components using IaC; maintain CI/CD pipelines; support monitoring, logging, and alerting; lead incident triage and root cause investigation; produce operational runbooks and deployment guides; participate in on-call rotation as L3 responder
Seniority
Mid-level, hands-on IC