Engineering Group Manager- MLOps
Core
Build and operate MLOps platforms on AWS supporting autonomous driving ML workloads, ensuring high availability, scalability, and safety-critical compliance for perception and decision-making systems.
Role type
Senior MLOps Platform Engineer (Autonomous Driving)
Builds
Highly available, multi-zone training and deployment environments; distributed multi-GPU clusters (Ray); production-ready ML pipelines for sensor-heavy datasets.
Domain
Automotive / Autonomous Driving / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
AWS (EKS, VPC, S3, Lambda, IoT), Kubernetes, Terraform, Apache Airflow, MLflow, Python, CI/CD (GitHub), distributed training (Ray), data validation/governance, monitoring/logging, incident response.
Preferred skills
Experience with perception models (vision, sensor fusion), automotive/ADAS environments, safety-critical system constraints.
Responsibilities
Build and operate MLOps platforms on AWS; implement multi-GPU training setups; maintain infrastructure standards for compute, storage, and security; support large-scale data ingestion and preprocessing; partner with ML engineers to improve training performance and reduce costs; build monitoring and logging for pipelines and deployed models.
Seniority
Senior, hands-on IC