Senior Manager, GPU Cloud Infrastructure - GeForce NOW
Core
Lead the design, scaling, and operations of high-performance networking for GPU-based cloud infrastructure to enable cloud gaming, AI/ML training, and inference platforms.
Role type
Senior Manager, Network Infrastructure
Builds
Ultra-low-latency, high-throughput interconnects across data centers and cloud environments for gaming and AI workloads
Domain
Cloud Gaming, AI/ML, Data Center Networking
Deliverable
production ML models | infrastructure
Required skills
Data center networking (Clos/spine-leaf, RDMA, RoCE, InfiniBand), BGP, EVPN/VXLAN, kernel-level routing/switching development, Infrastructure as Code (Ansible, Terraform), observability (Prometheus, Grafana), virtualization (SR-IOV, Xen, Open Virtual Switch), team leadership
Preferred skills
Large-scale GPU cluster networking, optical networking (400G/800G), Mellanox/Cumulus Linux debugging, Palo Alto/Netscaler management, streaming telemetry (SNMP, Syslog)
Technologies
RoCE, Ethernet-based AI fabrics, Ansible, Terraform, Prometheus, Grafana, SR-IOV, Xen, Open Virtual Switch, Mellanox, Cumulus Linux, Palo Alto, Netscaler
Responsibilities
Build and mentor a specialized team of network architects; oversee intra-cluster and inter-cluster connectivity design; drive technical tuning for latency, jitter, and throughput; define networking roadmaps for gaming and AI; engage with ISPs for edge network optimization; implement IaC and observability frameworks; establish fault tolerance protocols and lead incident response
Seniority
Senior, hands-on IC with management responsibilities