Site Reliability Engineer (Raptor)
Core
Manage server infrastructure, HPC systems, storage, and networks to accelerate rocket engine development and support engineering analysis software.
Role type
Site Reliability Engineer (Infrastructure & HPC)
Builds
High-performance computing clusters, storage systems, and networking infrastructure for propulsion engineering.
Domain
Aerospace / Rocket Propulsion / High Performance Computing
Deliverable
infrastructure
Required skills
Linux/Windows server management, enterprise networking, virtualization, security technologies, hardware/software troubleshooting
Preferred skills
Bash/Python scripting, automation (Puppet/Ansible), Kubernetes/Docker, CFD/FEA application support, performance bottleneck diagnosis
Technologies
Infiniband, ANSYS, StarCCM+, Puppet, Ansible, Kubernetes, Docker
Responsibilities
Manage server infrastructure, HPC systems, storage systems, networks, and high-speed interconnect; Design, procure, and integrate infrastructure systems; Work with propulsion engineering staff to solve critical bottlenecks; Support application deployment for best performance of software on real-world systems.
Seniority
Entry to Mid-level, hands-on IC