Senior Principal Network Engineer
Core
Design, deploy, and optimize next-generation AI data center networks for high-bandwidth, low-latency GPU clusters.
Role type
Senior Principal Network Engineer (AI Infrastructure)
Builds
High-performance computing (HPC) network fabrics supporting distributed AI training and inference workloads
Domain
AI compute infrastructure / Hyperscale Data Center Networking
Deliverable
production ML models | infrastructure
Required skills
Data center routing and switching protocols (BGP, OSPF, EVPN-VXLAN), RDMA networking (RoCEv2, InfiniBand), high-speed optics (400G/800G), automation/scripting (Python, Go, Bash), NetDevOps practices, network telemetry analysis, Clos spine-leaf-super-spine architecture design
Preferred skills
Large-scale AI/GPU cluster operations, network telemetry frameworks, vendor roadmap influence
Technologies
Arista EOS, Cisco NX-OS, SONiC, RoCEv2, InfiniBand, PFC, ECN, DCQCN
Responsibilities
Define ultra-high-bandwidth non-blocking AI network fabrics; optimize lossless Ethernet fabrics using congestion control; lead NetDevOps initiatives for automation; design high-resolution telemetry pipelines; model and deploy data center network fabrics; provide technical leadership and mentorship; contribute to long-term networking strategy; research next-generation networking technologies
Seniority
Principal, hands-on IC with strategic roadmap ownership
