Senior System Software Engineer - AI Data Platform - Inference Factory Optimization
Core
Design, build, and optimize highly scalable automation systems and infrastructure to ensure peak performance and seamless deployment of NVIDIA's core software offerings for AI and high-performance computing.
Role type
Senior System Software Engineer (Infrastructure & Performance)
Builds
Scalable automation systems, test harnesses, benchmarking frameworks, and analytical tools for software and hardware platforms.
Domain
High-performance computing, AI infrastructure, system software, distributed systems
Deliverable
production ML models | infrastructure
Required skills
System-level programming (C++, Python, Go), operating system internals, device drivers, memory management, distributed systems design, CI/CD pipeline development, performance profiling
Preferred skills
AI/ML workload optimization, containerization (Docker, Kubernetes), large-scale compute infrastructure experience
Responsibilities
Develop efficient infrastructure and tools for automating complex software processes; implement advanced test harnesses and benchmarking frameworks; build and troubleshoot highly performant systems; define requirements and deliver efficient solutions; monitor feedback and analyze data for continuous improvements; contribute to defining technical strategies and roadmaps.
Seniority
Senior, hands-on IC