Staff Senior Virtualization & Orchestration Engine
Core
Design and build virtualization infrastructure and Kubernetes-based orchestration systems for GPU-intensive AI and HPC workloads.
Role type
Staff Senior Infrastructure Engineer (GPU Orchestration)
Builds
Automated GPU cluster provisioning, workload scheduling, and multi-tenant resource management platforms.
Domain
Cloud Infrastructure / High-Performance Computing / AI Hardware
Deliverable
production ML models
Required skills
Kubernetes, container orchestration, Linux-based infrastructure, distributed systems, Infrastructure as Code, GPU cluster provisioning, workload scheduling, resource management, multi-tenant architecture.
Preferred skills
Slurm, Kubernetes device plugins, NVIDIA GPU Operator, high-performance networking, AI infrastructure platforms, HPC environments, GPU cloud providers, infrastructure automation, platform engineering.
Responsibilities
Design virtualization infrastructure for GPU workloads; develop Kubernetes orchestration for GPU clusters; build automated capacity allocation and scaling systems; design workload placement and cluster lifecycle solutions; partner on cluster architecture; improve platform reliability and security; build infrastructure tooling for developer experience; contribute to architectural decisions and engineering standards.
Seniority
Staff, hands-on IC with strategic impact