Senior Technical Program Manager (Engineering) - AI Tooling & Systems
Core
Drive execution of large-scale ML infrastructure and AI tooling initiatives, owning end-to-end delivery of programs spanning model serving, ML pipelines, and real-time inference systems.
Role type
Senior Technical Program Manager (ML Infrastructure & AI Tooling)
Builds
Model serving infrastructure, ML pipelines, internal AI tooling, and real-time inference systems
Domain
Artificial Intelligence / Machine Learning Infrastructure
Deliverable
production ML models
Required skills
End-to-end program delivery, technical architecture definition, cross-functional coordination, cost and latency optimization, bottleneck resolution, technical strategy alignment
Preferred skills
Model serving frameworks, LLM inference optimization, ML experiment tracking, feature stores, vector databases, multi-region deployment, cloud platforms
Technologies
vLLM, TensorRT, TorchServe, MLflow, Weights & Biases, DVC, PyTorch, CUDA, Hugging Face, AWS SageMaker, GCP Vertex AI, Azure ML
Responsibilities
Own end-to-end delivery of AI infrastructure programs; Define technical architecture and rollout strategies for new ML systems; Serve as connective tissue between research, engineering, and product teams; Drive cost and latency optimization for real-time inference; Build lightweight internal tools to accelerate ML iteration cycles; Identify and resolve technical bottlenecks in training and serving workflows; Translate research breakthroughs into scalable, observable systems
Seniority
Senior, hands-on IC