CareerPlanGet AI match score →

Principal AI Network HW Systems Engineer

United States, California, Mountain View💼 Full-time🗓 2026-07-20 → 2026-07-31

Core

Own AI networking architecture for MAIA training and inference systems across scale-up and scale-out deployments.

Role type

Principal AI Network HW Systems Engineer

Builds

AI networking architecture, switches, NICs, PHYs, optics, and high-speed SerDes technologies for MAIA systems.

Domain

AI infrastructure, high-speed networking, silicon photonics, hyperscale datacenter hardware.

Deliverable

production ML models | infrastructure

Required skills

AI accelerator system development (GPU/FPGA/TPU), high-speed interface architecture (SerDes), optical transceiver design (DR4/DR8/FR4), DSP architecture, optical link budget analysis, silicon photonics, IEEE Ethernet standards, OIF specifications, CMIS management frameworks, Linux environments, automation frameworks.

Preferred skills

collective communication efficiency optimization, latency and bandwidth utilization analysis, congestion management, large-scale distributed AI training environment experience, product ramp management (EVT/DVT/PVT).

Responsibilities

Lead integration and qualification of network components; define validation strategies for functionality, interoperability, and stress testing; optimize AI fabric performance; drive debugging and telemetry solutions.

Seniority

Principal, hands-on IC with strategy & mentorship

Rewrite
## About the role Own AI networking architecture for MAIA training and inference systems across scale-up and scale-out deployments. Lead integration and qualification of switches, NICs, PHYs, optics, cables, and high-speed SerDes technologies. Define validation strategies covering functionality, interoperability, performance, scale, reliability, and stress testing. Optimize AI fabric performance through analysis of latency, bandwidth utilization, congestion management, and collective communication efficiency. Drive debugging, telemetry, and automation solutions that improve network resiliency and operational excellence. ## Requirements * Master's Degree in Electrical Engineering, Computer Engineering, Mechanical Engineering, or related field AND 7+ years technical engineering experience * Bachelor's Degree in Electrical Engineering, Computer Engineering, Mechanical Engineering, or related field AND 8+ years technical engineering experience * Equivalent experience * 8+ years of experience in NW HW development * 8+ years of experience in GPU based SU/SO development * 8+ years of hands on experience with HS interface architecture and development * These requirements include but are not limited to the following specialized security screenings: * Experience developing GPU, FPGA, TPU, AI accelerator, or HPC-based systems. * Familiarity with AI workload communication patterns and collective operations. * Understanding of large-scale distributed AI training environments. * Optical transceivers (DR4, DR8, FR4) * DSP architectures TIAs and drivers * Optical link budgets * OMA, TDECQ, receiver sensitivity analysis * Silicon photonics technologies * Co-packaged optics architectures * IEEE Ethernet standards * OIF specifications * CMIS management frameworks * Ultra Ethernet Consortium technologies * Future 224G and 448G ecosystems * Experience with hyperscale datacenter deployments. * Experience managing products through EVT, DVT, PVT, and production ramps. * Experience with Linux environments and automation frameworks.
Sourced via microsoft · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply at Microsoft ↗