Distinguished Engineer, Storage – AI Cloud
Core
Lead the multi-year technical plan and architecture for high-performance parallel file systems, object stores, and block storage at exabyte scale to support NVIDIA's AI Cloud training and inference workloads.
Role type
Distinguished Engineer, Storage – AI Cloud
Builds
High-performance storage platforms (file, object, block) for AI training and inference across Neocloud and Cloud Service Provider ecosystems.
Domain
AI Infrastructure / High-Performance Storage / Cloud Computing
Deliverable
production ML models | infrastructure
Required skills
High-performance parallel file systems (Lustre, GPFS, DAOS), Object storage (S3, Swift), Block storage (NVMe-oF, iSCSI), Systems programming (C, C++, Rust, Go), Linux kernel storage stacks, Open-source strategy, AI coding tools
Preferred skills
Exabyte-scale storage management, Disaggregated inference architecture, KV caching, Cross-DC data movement, Public technical contributions (papers, patents)
Technologies
Lustre, GPFS, S3, NVMe-oF, SPDK, C++, Rust, Go, Python, Linux Kernel, RDMA, InfiniBand
Responsibilities
Define reference architecture and SLOs for NCP storage, Lead root cause analysis for complex production problems, Develop prototype reference implementations, Establish open-source strategy and APIs, Mentor senior engineers, Represent NVIDIA in industry standards bodies
Seniority
Distinguished, Strategy & Mentorship