HPC Systems Engineer
Core
Design, implement, maintain, and support high performance compute and storage systems for quantitative research data pipelines.
Role type
Senior HPC Systems Engineer
Builds
Production HPC infrastructure, monitoring tools, and software deployment tooling for research teams
Domain
High Performance Computing / Quantitative Finance
Deliverable
infrastructure
Required skills
Linux systems administration, parallel filesystems (Lustre, GPFS), batch systems (Slurm, Grid Engine), high-performance network interconnects, programming/scripting (Go, Python, C), distributed systems design, system profiling and debugging, configuration management (SaltStack, Ansible, Puppet), root cause analysis
Preferred skills
Experience with large-scale data storage integration
Responsibilities
Design and maintain HPC and storage systems; implement performance and fault monitoring; build software compilation and upgrade tooling; collaborate on testing infrastructures; develop system documentation; participate in coordinated maintenance operations; optimize HPC infrastructure usage for researchers; manage vendor relationships
Seniority
Senior, hands-on IC