Principal, Cloud Engineer
Core
Architect the long-term AWS cloud ecosystem, observability standards, and developer platform to ensure reliability, scalability, and security for a global healthcare platform.
Role type
Principal Site Reliability Engineer (Cloud Architecture & Platform)
Builds
Global AWS cloud infrastructure, unified observability ecosystems, internal developer platforms (IDP), and enterprise Kubernetes environments
Domain
Healthcare / Cloud Infrastructure / DevOps
Deliverable
infrastructure
Required skills
AWS ecosystem mastery, Kubernetes orchestration, Infrastructure-as-Code (Terraform/Terragrunt), Python/Go/Bash, observability architecture, security compliance (HIPAA/SOC 2), root-cause analysis, technical mentorship
Preferred skills
AWS Certified Solutions Architect – Professional, AWS Certified DevOps Engineer – Professional, distributed systems (Kafka/Kinesis), data warehousing (Redshift/Snowflake), open-source thought leadership
Technologies
AWS, Datadog, OpenTelemetry, GitHub Actions, ArgoCD, ACK, CUE, Kubernetes, Terraform, Terragrunt, Python, Go, Bash
Responsibilities
Define architectural strategy and technical governance for global AWS topologies; institutionalize observability standards using Datadog and OpenTelemetry; engineer CI/CD pipelines and internal developer platforms; steward enterprise Kubernetes clusters and service mesh architectures; architect zero-trust security and compliance controls; lead technical escalation and post-mortems for critical infrastructure failures; mentor senior engineers and define enterprise-wide best practices
Seniority
Principal, strategic architecture & mentorship