Senior Infrastructure Engineer - Infrastructure Security and Core Services
Core
Redesign core services for NVIDIA manufacturing sites, enabling new sites and troubleshooting incidents to ensure factory line uptime for GPU-accelerated OT and IT environments.
Role type
Senior Infrastructure Engineer (Manufacturing Security & Core Services)
Builds
Global-scale manufacturing site connectivity, CPU-based compute/storage/GPU-HPC clusters, high-performance OT/IT networks, and secure infrastructure for AI factories.
Domain
Semiconductor manufacturing, High-Performance Computing (HPC), Artificial Intelligence infrastructure
Deliverable
production ML models | infrastructure
Required skills
Large-scale hybrid network design, Infrastructure automation (Python, Ruby, Go), Networking technologies (Mellanox, routing, switching), Compute/Storage architecture (Dell, Pure, NetApp, Openshift), Data center design, Test automation infrastructure, Scripting, DNS/DHCP troubleshooting, Security & compliance standards
Preferred skills
Experience with factory floor infrastructure, Test Engineering topologies, Designing CPU and GPU workloads, Cloud-based infrastructure services
Technologies
Mellanox, Dell, Openshift, Pure, NetApp, Python, Ruby, Go
Responsibilities
Lead architecture and deployment of global manufacturing site connectivity; Design high-performance OT/IT networks for general compute and GPU-dense AI/ML; Partner with product teams to deliver scalable network architectures; Implement compute, storage, security, and performance engineering practices; Manage lifecycle of revenue-generating manufacturing sites; Enforce security and reliability standards for mission-critical workloads; Troubleshoot full connectivity stack infrastructures; Support infrastructure on the factory floor.
Seniority
Senior, hands-on IC
