Senior Site Reliability Engineer – AUS Region
Core
Maintain reliable, secure, and AI-assisted production operations for a vulnerability management platform serving large enterprise organizations.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Kubernetes clusters, containerized workloads, Infrastructure as Code, CI/CD pipelines, and cloud infrastructure
Domain
Cyber security / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, cloud platforms (AWS, GCP, Azure, OpenShift), Infrastructure as Code (Terraform), automation scripting (Python, Bash), observability (Prometheus, Grafana, Loki), incident response, Linux, networking, distributed systems
Preferred skills
LLM integration, AI-assisted log analysis, anomaly detection, internal automation tooling combining APIs and AI models
Technologies
Kubernetes, AWS, GCP, Azure, OpenShift, Terraform, Ansible, Python, Bash, GitHub, GitLab, Bitbucket, Prometheus, Grafana, Loki, CloudWatch
Responsibilities
Maintain highly available, secure, and patched production systems; Build and improve Kubernetes clusters and cloud infrastructure; Improve monitoring, alerting, logging, and automated remediation workflows
Seniority
Senior, hands-on IC with technical leadership