Principal Cloud & Kubernetes Engineer
Core
Principal Kubernetes Engineer leading the build, operation, automation, and continuous improvement of mission-critical healthcare and financial services platforms on Kubernetes-based managed services and SaaS.
Role type
Principal, hands-on IC
Builds
Resilient, automated, and scalable Kubernetes-based managed services and SaaS platforms for healthcare and financial services clients.
Domain
Cloud-native infrastructure, Kubernetes, Site Reliability Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Kubernetes (production), Cloud platforms (AWS/Azure/GCP), Infrastructure as Code (Terraform), GitOps (Argo CD/Flux), Linux, Container runtimes, Networking, Observability (Prometheus/Grafana), Security (RBAC/Network policies), Automation scripting (Python/Bash/Go)
Preferred skills
Multi-cloud/hybrid experience, Mentorship of engineers, Incident management
Technologies
Kubernetes, EKS, AKS, GKE, Rancher, Spectro Cloud, Terraform, Helm, Argo CD, Flux, Prometheus, Grafana, Loki, Fluent Bit, CloudWatch, Azure Monitor
Responsibilities
Lead complex Kubernetes platform engineering initiatives across public, private, and on-premises environments; Translate reference architectures into production-ready implementation patterns; Own advanced cluster lifecycle activities including provisioning, upgrades, patching, and resource governance; Act as a senior technical escalation point for platform design, operations, and incident resolution; Lead troubleshooting of complex Kubernetes failures involving control plane, node health, CNI, ingress, and storage; Drive root cause analysis and permanent corrective actions for recurring platform issues; Mentor Kubernetes Engineers and Senior Engineers through design reviews, code reviews, and operational coaching.
Seniority
Principal, hands-on IC with mentorship responsibilities