CareerPlanSign in

Head of Supercomputing

San Jose💼 Full-time🗓 2026-03-26 → 2026-09-25

Core

Define and lead the architecture, software stack, and operational model for cluster-scale AI compute systems, owning end-to-end system software and control-plane strategy from first silicon through production deployment.

Role type

Head of Supercomputing (Senior IC + Engineering Leadership)

Builds

Cluster-scale AI inference systems, control-plane software, orchestration primitives, and fleet management tools

Domain

AI Infrastructure / High-Performance Computing (HPC) / Custom Accelerators

Deliverable

production ML models | infrastructure

Required skills

System software architecture, cluster-scale system operations, hardware/software interface expertise, engineering team leadership, low-level control-plane development, system telemetry and observability, cross-functional collaboration

Preferred skills

Experience with ASICs, kernel subsystems, firmware integration, manufacturing test engineering

Technologies

PCIe, RDMA, memory hierarchies, interrupts, device drivers, firmware, kernel subsystems, runtime layers

Responsibilities

Define technical vision and roadmap for supercomputing software stack, architect low-level control-plane software, manage and develop 15+ engineers, oversee system services interfacing with hardware/firmware, establish telemetry infrastructure, define reliability targets and release processes, recruit and mentor engineering leaders

Seniority

Principal, hands-on IC with strategic leadership

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.