Site Reliability Engineer
Core
Empower the Content Engineering team by providing reliable, scalable, and automated cloud infrastructure for hands-on cybersecurity learning experiences.
Role type
Site Reliability Engineer (Infrastructure & Automation)
Builds
Cloud labs, production services, and automated workflows for content creation and delivery
Domain
Cybersecurity education platform / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Terraform, Infrastructure as Code, Google Cloud Platform, Microsoft Azure, scripting (Go/Python/Bash), observability, CI/CD pipelines
Preferred skills
Kubernetes, containers, cloud-native technologies, incident response, application development workflows
Technologies
Terraform, Google Cloud Platform, Microsoft Azure, Prometheus, Grafana, Mimir, Loki, Tempo
Responsibilities
Automate provisioning and management of cloud resources, maintain visibility into platform health, improve developer experience and platform adoption, support production environments through maintenance and troubleshooting, drive automation initiatives to reduce manual effort
Seniority
Mid-level, hands-on IC