High Performance Computing Engineer
Core
Provisioning, configuration, and lifecycle management of secure and scalable HPC environments for computational research.
Role type
HPC Infrastructure Engineer
Builds
Compute clusters, workload scheduling systems, and secure user-facing software environments
Domain
Research Computing / High-Performance Computing
Deliverable
infrastructure
Required skills
Linux system administration, workload scheduler management (Slurm), cluster provisioning, performance tuning, infrastructure monitoring, configuration management (Ansible), containerization (Apptainer/Singularity, Docker), security compliance (NIST 800-171), scripting and automation
Preferred skills
Experience in research or academic environments, Harvard IT Academy foundational courses
Technologies
Slurm, Ansible, Apptainer, Singularity, Docker, NIST 800-171
Responsibilities
Provision, configure, and decommission HPC compute clusters; Administer and tune workload schedulers; Maintain secure, regulated compute environments; Integrate user accounts and identity management; Maintain and optimize user-facing software environments; Develop and maintain automation scripts; Monitor system health and respond to alerts; Collaborate with researchers to troubleshoot and improve the environment
Seniority
Mid-level, hands-on IC