Staff Technical Program Manager, Managed Intelligence
Core
Own end-to-end program delivery for the Managed Inference platform, connecting model engineering, IaaS, and data center operations to deliver reliable, scalable LLM workloads.
Role type
Staff Technical Program Manager (AI Infrastructure)
Builds
Production-ready Managed Inference platform for AI-native companies
Domain
AI Infrastructure / Cloud Computing / Data Centers
Deliverable
production ML models
Required skills
End-to-end program delivery, LLM inference and model serving knowledge, Multi-tenant systems experience, Fine-tuning and alignment awareness, Low-structure execution, Executive communication, Cross-functional influence
Preferred skills
Experience governing model onboarding programs across GPU generations, Coaching or mentoring junior TPMs, Background at a Series D to Series F company or hyperscaler
Technologies
CUDA, ROCm, GPU generations
Responsibilities
Own multi-quarter release planning and dependency governance, Drive model version rollouts and inference optimization campaigns, Coordinate across Model Engineering, IaaS, and Data Center Operations, Surface risks across model serving and capacity constraints, Build lightweight execution frameworks and dashboards, Own pre-launch planning for model onboarding including firmware and driver validation
Seniority
Staff, hands-on IC with strategic scope