CareerPlanSign in

Cloud Engineer ( Azure Platform Engineer)- Remote Ontario

Toronto, Ontario🌐 Remote💼 Full-time🗓 2026-07-30 → 2026-09-26

Core

Design, deploy, and operate production-grade Azure Kubernetes Service (AKS) clusters, focusing on cluster internals, networking, security hardening, and reliability for healthcare data platforms.

Role type

Senior IC Azure Platform Engineer (Kubernetes)

Builds

Production AKS clusters, CI/CD pipelines, observability stacks, and disaster recovery procedures for healthcare data systems.

Domain

Healthcare technology / Cloud Infrastructure

Deliverable

production ML models | infrastructure

Required skills

Kubernetes internals, Azure CNI networking, AKS private cluster configuration, Helm, GitOps (Flux/ArgoCD), Terraform, Azure DevOps, Grafana stack (Prometheus, Loki, Tempo), Azure-native monitoring, container vulnerability management, Bash/Python scripting

Preferred skills

None stated

Technologies

Azure Kubernetes Service (AKS), Azure Container Registry (ACR), Azure Key Vault, Azure SQL, Kafka, Event Hubs, Azure Storage, NGINX, Traefik, Docker, Helm, Terraform, Azure DevOps, Flux, ArgoCD, Grafana, Prometheus, Loki, Tempo, Azure Monitor, Log Analytics, Application Insights

Responsibilities

Act as SME for Kubernetes deployments, troubleshooting, and production issues; Design and deploy AKS clusters with private configurations, managed identities, and RBAC; Own health, scaling, and lifecycle management of production AKS clusters; Configure AKS networking including Azure CNI, internal load balancers, and ingress controllers; Design and maintain integrations between AKS and Azure PaaS services; Manage containerized application deployments using Docker and Helm; Harden AKS environments through policy enforcement and image scanning; Own container and cluster vulnerability management; Contribute to Terraform-based infrastructure as code; Support Azure DevOps CI/CD pipelines including GitOps workflows; Design, implement, and manage observability across Grafana stack and Azure-native tooling; Collaborate with performance engineering to identify and resolve performance bottlenecks; Contribute to Disaster Recovery and Business Continuity Planning; Provide escalation support for production incidents and lead root-cause analysis; Document runbooks and post-incident reviews

Seniority

Senior, hands-on IC

Sourced via lever · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.