CareerPlanSign in

Supercomputing Engineer

San Jose💼 Full-time🗓 2025-06-18 → 2026-09-25

Core

Develop foundational software for cluster-scale AI compute deployments, focusing on control-plane systems, hardware integration, and performance tuning.

Role type

Senior IC systems engineer (AI infrastructure)

Builds

Control-plane software, system services, telemetry infrastructure, and orchestration primitives for AI clusters

Domain

AI infrastructure / High-performance computing

Deliverable

production ML models

Required skills

C/C++ or Rust, Linux kernel internals, hardware driver development, system-level debugging, PCIe/memory/networking tuning

Preferred skills

Kubernetes/Docker, eBPF/perf/ftrace, HPC background, failure-injection frameworks

Technologies

C, C++, Rust, Linux, PCIe, eBPF, perf, ftrace, Kubernetes, Docker

Responsibilities

Architect low-level control-plane software for system bring-up and management; Build system services interacting with hardware/firmware/OS; Develop telemetry and tracing infrastructure; Implement orchestration primitives for nodes and racks; Profile and tune performance across kernel and runtime layers; Collaborate with hardware and firmware teams on system interfaces

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.