Cloud Engineer ( Azure Platform Engineer)- Remote Ontario
Core
Design, deploy, and operate production-grade Azure Kubernetes Service (AKS) clusters, focusing on cluster internals, networking, security hardening, and reliability for healthcare data platforms.
Role type
Senior IC Azure Platform Engineer (Kubernetes)
Builds
Production AKS clusters, CI/CD pipelines, observability stacks, and disaster recovery procedures for healthcare data systems.
Domain
Healthcare technology / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes internals, Azure CNI networking, AKS private cluster configuration, Helm, GitOps (Flux/ArgoCD), Terraform, Azure DevOps, Grafana stack (Prometheus, Loki, Tempo), Azure-native monitoring, container vulnerability management, Bash/Python scripting
Preferred skills
None stated
Technologies
Azure Kubernetes Service (AKS), Azure Container Registry (ACR), Azure Key Vault, Azure SQL, Kafka, Event Hubs, Azure Storage, NGINX, Traefik, Docker, Helm, Terraform, Azure DevOps, Flux, ArgoCD, Grafana, Prometheus, Loki, Tempo, Azure Monitor, Log Analytics, Application Insights
Responsibilities
Act as SME for Kubernetes deployments, troubleshooting, and production issues; Design and deploy AKS clusters with private configurations, managed identities, and RBAC; Own health, scaling, and lifecycle management of production AKS clusters; Configure AKS networking including Azure CNI, internal load balancers, and ingress controllers; Design and maintain integrations between AKS and Azure PaaS services; Manage containerized application deployments using Docker and Helm; Harden AKS environments through policy enforcement and image scanning; Own container and cluster vulnerability management; Contribute to Terraform-based infrastructure as code; Support Azure DevOps CI/CD pipelines including GitOps workflows; Design, implement, and manage observability across Grafana stack and Azure-native tooling; Collaborate with performance engineering to identify and resolve performance bottlenecks; Contribute to Disaster Recovery and Business Continuity Planning; Provide escalation support for production incidents and lead root-cause analysis; Document runbooks and post-incident reviews
Seniority
Senior, hands-on IC