CareerPlanSign in

Principal Product Manager/Architect - Foundry Inference Platform (CoreAI)

United States, Washington, Redmond💼 Full-time🗓 2026-03-24 → 2026-09-26

Core

Define architectural standards for global serving, multi-region resiliency, and platform-managed disaster recovery for the Foundry Inference Platform while driving GPU fleet efficiency and capacity management.

Role type

Principal Product Manager/Architect (Inference Platform)

Builds

Global inference serving platform, GPU capacity pooling, and automated scheduling primitives

Domain

Cloud AI / High-performance compute / GPU inference

Deliverable

production ML models

Required skills

Architecting planet-scale distributed systems, designing platform-managed resilience and disaster recovery, GPU-backed inference system design, capacity management and scheduling, strategic customer engagement at CTO level, translating complex technical concepts for executives

Preferred skills

Experience with cloud AI or high-performance compute platforms, owning end-to-end architecture for mission-critical services, deep understanding of model serving and hardware/system performance optimization

Technologies

GPU-backed inference systems, global capacity pooling, automated demand forecasting, software-defined allocation

Responsibilities

Define architectural standards for global serving and multi-region resiliency, set product direction for GPU fleet efficiency and capacity management, act as a senior technical advisor for strategic customers on large-scale model migrations, drive alignment on long-term technical direction across product, engineering, and infrastructure teams, improve revenue per GPU and time to revenue through systemic efficiency gains

Seniority

Principal, strategy & mentorship

Sourced via microsoft · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.