System Engineer (Compute Node)
Core
Building services for managing Virtual Machines on GPU servers, integrating disk management, virtual and InfiniBand networks, and developing a Virtual Machine Scheduler for clusters with thousands of servers and GPUs.
Role type
Senior IC system engineer (compute infrastructure)
Builds
Virtual Machine Scheduler, GPU server management services, cluster orchestration tools
Domain
Cloud infrastructure, AI compute, hardware acceleration
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Linux system architecture, Go, C++, virtualization (KubeVirt/QEMU), concurrency, debugging, profiling
Preferred skills
GPU/DPU/ARM architecture, NVIDIA DOCA Software Framework, InfiniBand networking
Technologies
Kubernetes, KubeVirt, QEMU/KVM, InfiniBand, NVIDIA DOCA, Go, C++
Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.