Principal Product Manager/Architect - Foundry Inference Platform (CoreAI)
Core
Define architectural standards for global serving, multi-region resiliency, and platform-managed disaster recovery for the Foundry Inference Platform while driving GPU fleet efficiency and capacity management.
Role type
Principal Product Manager/Architect (Inference Platform)
Builds
Global inference serving platform, GPU capacity pooling, and automated scheduling primitives
Domain
Cloud AI / High-performance compute / GPU inference
Deliverable
production ML models
Required skills
Architecting planet-scale distributed systems, designing platform-managed resilience and disaster recovery, GPU-backed inference system design, capacity management and scheduling, strategic customer engagement at CTO level, translating complex technical concepts for executives
Preferred skills
Experience with cloud AI or high-performance compute platforms, owning end-to-end architecture for mission-critical services, deep understanding of model serving and hardware/system performance optimization
Technologies
GPU-backed inference systems, global capacity pooling, automated demand forecasting, software-defined allocation
Responsibilities
Define architectural standards for global serving and multi-region resiliency, set product direction for GPU fleet efficiency and capacity management, act as a senior technical advisor for strategic customers on large-scale model migrations, drive alignment on long-term technical direction across product, engineering, and infrastructure teams, improve revenue per GPU and time to revenue through systemic efficiency gains
Seniority
Principal, strategy & mentorship
