Senior HPC Architect, Automation and At-Scale Deployment
Core
Architect and engineer large-scale GPU compute clusters, implementing at-scale system administration and tuning for high-performance computing and deep learning workloads.
Role type
Senior HPC Architect (Infrastructure & Systems)
Builds
Large-scale GPU-accelerated compute platforms and optimized workflows for scientific research and enterprise AI.
Domain
High-Performance Computing (HPC), Accelerated Computing, Datacenter Infrastructure
Deliverable
infrastructure
Required skills
Accelerated computing scheduling, I/O stacks, C/C++/Python/Bash scripting, parallel filesystems, system administration, performance tuning
Preferred skills
Deep Learning frameworks, telemetry and visualization pipelines, container technology, Linux performance tools
Technologies
GPU compute platforms, parallel filesystems, container technology, Linux
Responsibilities
Provide engineering solutions to operationalize GPU computing products and software stacks; act as an internal reference for system administration and at-scale system analysis; architect, develop, and bring up large-scale performance platforms with HPC and systems specialists.
Seniority
Senior, hands-on IC