Compute Deployment Engineer
Core
Deploy and qualify large-scale GPU and custom accelerator fleets from facility readiness to production, automating hardware workflows and managing remote/on-site turn-up operations.
Role type
Senior IC compute deployment engineer (hardware/software)
Builds
Production-ready compute clusters for AI workloads
Domain
AI infrastructure, data centers, hardware provisioning
Deliverable
production ML models | infrastructure
Required skills
Linux administration, out-of-band management (BMC, IPMI, Redfish), Python/Go automation, hardware failure triage, rack-level qualification, remote hands management
Preferred skills
Kubernetes bare-metal provisioning, accelerator platform bringup (NVIDIA/AMD), burn-in/stress harness design, DCIM tooling
Technologies
Kubernetes, Python, Go, BMC, IPMI, Redfish
Responsibilities
Own compute turn-up from facility availability to ready-for-service; Qualify racks at scale via firmware baselines, BMC/BIOS config, and burn-in; Drive qualification through base-management Kubernetes platform; Triage hardware failures and feed patterns back into qual gates; Run remote turn-up with periodic on-site pulses; Partner with network, ICT, and hardware teams during turn-up windows
Seniority
Senior, hands-on IC