Supercomputing Engineer
Core
Develop foundational software for cluster-scale AI compute deployments, focusing on control-plane systems, hardware integration, and performance tuning.
Role type
Senior IC systems engineer (AI infrastructure)
Builds
Control-plane software, system services, telemetry infrastructure, and orchestration primitives for AI clusters
Domain
AI infrastructure / High-performance computing
Deliverable
production ML models
Required skills
C/C++ or Rust, Linux kernel internals, hardware driver development, system-level debugging, PCIe/memory/networking tuning
Preferred skills
Kubernetes/Docker, eBPF/perf/ftrace, HPC background, failure-injection frameworks
Technologies
C, C++, Rust, Linux, PCIe, eBPF, perf, ftrace, Kubernetes, Docker
Responsibilities
Architect low-level control-plane software for system bring-up and management; Build system services interacting with hardware/firmware/OS; Develop telemetry and tracing infrastructure; Implement orchestration primitives for nodes and racks; Profile and tune performance across kernel and runtime layers; Collaborate with hardware and firmware teams on system interfaces
Seniority
Senior, hands-on IC