Staff ML Engineer, ML Compute Platform
Core
Design and implement core backend software components for a cloud-agnostic compute platform that powers ML workflows for autonomous vehicles and AI-driven products.
Role type
Staff ML Engineer (Infrastructure)
Builds
Cloud-agnostic, reliable, and cost-efficient compute backend for training and deploying SOTA machine learning models.
Domain
Automotive AI / Cloud Infrastructure
Deliverable
infrastructure
Required skills
Go, C++, Python, Kubernetes at scale, distributed systems, cloud platforms (GCP, Azure, AWS), technical leadership
Preferred skills
ML infrastructure platforms, job orchestration interfaces, observability, GPU/TPU optimizations, PyTorch, TorchX, Ray framework, open source contributions
Technologies
Kubernetes, GCP, Azure, AWS, PyTorch, TorchX, Ray, B200, H100, A100
Responsibilities
Design and implement core platform backend software components; Collaborate with ML engineers to improve developer experience; Analyze and improve efficiency, scalability, and stability of system resources; Lead large-scale technical initiatives across the ML ecosystem; Contribute to and lead open source projects.
Seniority
Staff, hands-on IC with leadership