NVIDIA - AI Infrastructure Specialist (GDC) - 62808
Core
Deploying and managing Kubernetes clusters for AI/ML workloads at scale, utilizing NVIDIA AI Enterprise components and infrastructure tools.
Role type
Senior IC AI Infrastructure Specialist
Builds
AI/ML workloads on Kubernetes clusters
Domain
Cloud infrastructure, AI/ML, NVIDIA ecosystem
Deliverable
production ML models
Required skills
Kubernetes cluster management, Azure Kubernetes Service, RedHat OpenShift, Microk8s, Helm Charts, infrastructure and resource management, virtualization tools (VMWare/EXSi, KVM, Ansible, Redfish), Run:AI platform, NVIDIA AI Enterprise components (NIM, NeMO, TAO, Triton, Nucleus Servers), DGX systems, Jetson, NVIDIA AI Factory, Python, C++
Preferred skills
.NET/C#
Technologies
Kubernetes, Azure Kubernetes Service, RedHat OpenShift, Microk8s, Helm Charts, VMWare, EXSi, KVM, Ansible, Redfish, Run:AI, NVIDIA AI Enterprise, DGX, Jetson
Responsibilities
Deploy and manage Kubernetes clusters for AI/ML workloads, manage at-scale deployments with Azure Kubernetes Service, utilize virtualization tools for infrastructure management, configure Run:AI platform for job scheduling and GPU virtualization, integrate NVIDIA AI Enterprise components
Seniority
Senior, hands-on IC