Staff Platform Engineer
Core
Define and lead platform engineering strategy for complex, multi-environment cloud systems supporting AI/ML workloads and enterprise operations.
Role type
Staff Platform Engineer (Infrastructure & AI/ML)
Builds
Scalable Kubernetes platforms, AI/ML infrastructure (model serving, GPU orchestration, vector stores), and production-ready cloud systems.
Domain
Cloud Infrastructure, DevOps, AI/ML Platform Engineering
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Infrastructure as Code, CI/CD, DevSecOps, Cloud Architecture, Python/Go/Java/Bash, AI/ML Platform Infrastructure, Zero-trust networking, Service mesh, Distributed systems
Preferred skills
Cloud certifications (AWS/Azure), FinOps, AI-forward tool usage (Claude, Cursor)
Technologies
Kubernetes, AWS, Azure, Python, Go, Java, Bash, Terraform, Helm, Prometheus, Grafana, Docker
Responsibilities
Define DevOps strategy and lead infrastructure architecture; Architect and own scalable Kubernetes platforms; Own infrastructure as code strategy; Lead DevSecOps implementation; Drive platform reliability, performance SLAs, and cost optimization; Lead complex cloud migrations and platform modernization; Own observability strategy; Lead design and operation of AI/ML platform infrastructure; Mentor engineers and partner with leadership on infrastructure direction.
Seniority
Staff, hands-on IC with leadership & mentorship

