Senior Specialist Field Engineer - Compute Infrastructure
Core
Own the technical path for large-scale GPU cluster delivery, ensuring seamless, reliable, high-performance compute infrastructure for AI workloads.
Role type
Senior Specialist Field Engineer - Compute Infrastructure
Builds
Production-ready supercomputers, bare-metal fleets, and high-performance compute (HPC) environments for AI labs and enterprises.
Domain
Cloud Infrastructure / High-Performance Computing (HPC) / AI Compute
Deliverable
production ML models | infrastructure
Required skills
Bare-metal compute infrastructure expertise, large-scale GPU cluster delivery, InfiniBand/RoCE fabric validation, HPC performance benchmarking, Linux system administration, networking fundamentals (routing, fabric topologies, TCP/IP), rack-scale GPU server hardware knowledge, Kubernetes/Slurm integration
Preferred skills
Security-sensitive/air-gapped environment operations, infrastructure automation scripting (Python, Bash, Ansible), MEP design for AI supercomputers, multi-cloud/hybrid environment solutions
Technologies
InfiniBand, RoCE, NVIDIA HGX, GB200, Kubernetes, Slurm, Python, Bash, Ansible
Responsibilities
Lead bring-up and acceptance of new large-scale GPU clusters; drive fabric validation and HPC performance benchmarking; define and operationalize models for managing customer bare-metal fleets; partner with operations teams to align readiness with go-live timelines; review and advise on technical contract terms; lead proof of concept initiatives; represent CoreWeave at industry events.
Seniority
Senior, hands-on IC