Senior HPC Cluster Engineer
Core
Design, build, and maintain internal high-performance computing (HPC) platforms spanning on-premises and cloud clusters to support quantum computing research and simulations.
Role type
Senior HPC Cluster Engineer (Infrastructure)
Builds
Scalable HPC infrastructure for quantum simulations and applications
Domain
Quantum Computing / High-Performance Computing
Deliverable
production ML models | infrastructure
Required skills
Linux cluster management, HPC workload scheduling (Slurm/PBS/Grid Engine), Infrastructure automation (Ansible), Scripting (Python/Go/Bash), Version control (Git), High-performance networking (InfiniBand/RoCE), Containerization (Docker/Podman/Singularity/Enroot), Cloud platforms (AWS/GCP)
Preferred skills
NVIDIA GPU cluster deployment, Kubernetes for long-running workloads, Quantum physics/computing background, Scientific computing experience
Technologies
Slurm, PBS, Grid Engine, Ansible, Python, Go, Bash, Git, InfiniBand, RoCE, Docker, Podman, Singularity, Enroot, Kubernetes, AWS, GCP, NVIDIA GPUs
Responsibilities
Design, deploy, and operate HPC systems for quantum simulations; Develop automation tooling for Linux system configuration; Support researchers and developers running workloads; Build technical relationships with stakeholders; Self-manage projects and mitigate technical risks
Seniority
Senior, hands-on IC
