AI Infrastructure Associate Engineer (3 x FTE)
Core
Design and operate large, highly available supercomputing services and compute platforms for researchers conducting leading-edge AI research.
Role type
Associate Engineer, AI Infrastructure (Storage & Networking focus)
Builds
Software-defined supercomputing infrastructure, GPU/CPU workloads, and research platforms
Domain
High-performance computing (HPC), AI research infrastructure, supercomputing
Deliverable
infrastructure
Required skills
Python, Rust, Terraform/OpenTofu, Kubernetes, Git, Bash, cluster operations, distributed systems management
Preferred skills
SysOps, NetOps, DevOps, SecOps, MLOps, Research Software Engineering
Technologies
Kubernetes, Terraform, OpenTofu, Python, Rust, Git, Bash
Responsibilities
Design and operate massive-scale GPU and combined CPU/GPU workloads; design and debug platforms for computational experiments; co-design solutions with researchers to enable new algorithms and software; maintain and secure software-defined supercomputing systems
Seniority
Associate Engineer, hands-on IC