INFRASTRUCTURE & HPC SYSTEMS ENGINEER
Core
Manage Windows/Linux server infrastructure, HPC clusters, and cloud/on-prem services to ensure integrity and availability of agile research computing environments for economists and researchers.
Role type
Senior Infrastructure & HPC Systems Engineer
Builds
High-performance computing clusters, automated management tools, and reproducible research environments
Domain
Financial services / High-Performance Computing / Research Computing
Deliverable
infrastructure
Required skills
Linux administration (Red Hat/CentOS), HPC cluster design and administration, job scheduling (SLURM), automation scripting (Python, Bash), parallel file systems (Ceph, GPFS, Lustre), containerization (Docker), configuration management (Terraform)
Preferred skills
GPU computing support, parallel programming models (MPI, OpenMP), scientific computing frameworks, vendor management
Technologies
Slurm, Docker, Terraform, Ceph, GPFS, Lustre, BeeGFS, MPI, OpenMP
Responsibilities
Design, deploy, and administer HPC clusters; monitor system health and resource utilization; develop automation scripts to streamline system administration; troubleshoot complex hardware and software issues in multi-user research environments; create user guides and conduct training sessions on HPC resources
Seniority
Senior, hands-on IC