Group Product Manager, Compute Platform
Core
Define and own the foundational compute platform (fleet, capacity, bare metal, virtualized) powering enterprise AI workloads, turning physical resources into reliable, billable products.
Role type
Group Product Manager, Compute Infrastructure
Builds
Next-generation cloud platform for accelerated computing, including fleet management, capacity management, and node lifecycle services.
Domain
Cloud Infrastructure / High-Performance Computing (HPC) / AI Workloads
Deliverable
production ML models | product features
Required skills
Product strategy & roadmap definition, Infrastructure capacity management, Enterprise customer requirements translation, Cross-functional partnership (engineering, sales, finance), Compute abstraction design, Lifecycle & observability management
Preferred skills
Experience with Kubernetes, Slurm, Ray, or similar orchestration systems, Understanding of AI training and inference workloads
Technologies
Kubernetes, Slurm, Ray, Bare Metal, Virtualized Compute
Responsibilities
Define product strategy and roadmap for compute platform across fleet, capacity, bare metal, and virtualized compute; Own product lifecycle from infrastructure capacity to customer-ready compute including reservation, provisioning, and billing; Define customer-facing abstractions for capacity, node pools, and reserved compute; Partner with engineering to define fleet readiness and operational requirements; Partner with sales and finance to translate enterprise AI workload requirements into compute offerings