Site Reliability Engineer (NZ)
Core
Ensure reliability, availability, and performance of critical customer-facing AI platforms for cancer detection.
Role type
Intermediate Site Reliability Engineer
Builds
Azure-based cloud infrastructure, CI/CD pipelines, and automation tooling
Domain
Healthcare AI / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Azure cloud management, CI/CD pipeline development, Infrastructure as Code (Bicep/ARM/Terraform), PowerShell/Bash scripting, production incident response, observability practices, Git, SQL
Preferred skills
Linux administration, Azure Entra ID, Docker/container technologies, ISO/IEC 27001, ServiceNow/ITIL, Snowflake/Dagster, C# development
Technologies
Azure, GitHub, Azure DevOps, Bicep, ARM, Terraform, PowerShell, Bash, SQL, Docker, C#, ServiceNow, Snowflake, Dagster
Responsibilities
Support and enhance Azure-based cloud infrastructure; Develop and maintain CI/CD pipelines; Build and improve Infrastructure as Code solutions; Create tooling and automation to reduce operational overhead; Improve monitoring, alerting, logging, and observability; Investigate production incidents through to resolution; Lead blameless post-incident reviews; Participate in an on-call roster (approx. one week per month)
Seniority
Intermediate, hands-on IC