HPC Systems Engineer
Core
Design, implement, maintain, and support high performance compute and storage systems for quantitative research data pipelines.
Role type
Senior HPC Systems Engineer
Builds
Production HPC infrastructure, monitoring tools, and software deployment tooling for research teams.
Domain
Financial technology / High Performance Computing
Deliverable
infrastructure
Required skills
Linux systems administration, HPC cluster management, parallel filesystems, batch scheduling systems, high-performance networking, system configuration management, software development (Go/Python/C), performance profiling and debugging, root cause analysis
Preferred skills
Lustre, GPFS, Slurm, Grid Engine, SaltStack, Ansible, Puppet
Technologies
Linux, Go, Python, C, Lustre, GPFS, Slurm, Grid Engine, SaltStack, Ansible, Puppet
Responsibilities
Design and maintain HPC and storage systems; implement performance and fault monitoring; build tooling for software compilation and OS upgrades; collaborate on testing infrastructures; develop system documentation; manage vendor relationships; provide operational support; optimize HPC usage for researchers.
Seniority
Senior, hands-on IC