Staff SRE SW Development Engineer
Core
Architect, build, and operate resilient, scalable, and secure cloud infrastructure for Dexcom's R&D Platform serving millions of customers daily.
Role type
Staff Site Reliability Engineer (Cloud Infrastructure & Kubernetes)
Builds
Cloud-native R&D platform, Kubernetes clusters, observability ecosystem, and CI/CD pipelines on Google Cloud Platform.
Domain
Consumer health technology / Continuous Glucose Monitoring (CGM) / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes operational mastery, Google Cloud Platform (GCP) expertise, Infrastructure as Code (Terraform/Pulumi/Crossplane), Python/Go/Bash scripting, SLO/SLA framework design, incident management, root cause analysis, automation strategy.
Preferred skills
HIPAA/ISO/medical device compliance experience, CKA/CKAD/GCP Professional Cloud Engineer certifications, multi-tenant architecture design.
Technologies
Google Cloud Platform (GCP), Kubernetes, Terraform, Pulumi, Crossplane, Python, Go, Bash
Responsibilities
Architect and evolve observability ecosystem; Design and operate highly available cloud infrastructure; Lead Kubernetes platform operations; Diagnose and resolve complex infrastructure failures; Set direction for Infrastructure as Code; Drive automation strategy; Lead major incident response and post-incident reviews; Mentor engineers on operational discipline.
Seniority
Staff, strategic leadership with hands-on technical execution