Manager, System Software Engineering
Core
Lead development of an efficient on-device AI software stack for high-performance local inference, agentic workloads, and low-latency execution on RTX, RTX Pro, and DGX-class systems.
Role type
Senior IC system software manager (on-device AI inference)
Builds
On-device AI inference platform and execution stacks for RTX, RTX Pro, and DGX GPUs
Domain
AI computing, GPU systems, local inference
Deliverable
production ML models
Required skills
C++ software development, AI inference pipelines, Deep Learning frameworks, inference backends and runtime internals, model optimization techniques, team leadership, cross-functional alignment
Preferred skills
CUDA programming, high-performance systems development, open-source contributions, distributed team management
Technologies
Llama.cpp, vLLM, PyTorch, WinML, DXCGC, TensorRT-RTX, Vulkan, DirectX
Responsibilities
Lead and grow a team building the on-device AI inference platform; Drive cross-functional alignment with software, research, architecture, and product teams; Provide technical leadership for architecture and evolution of inference runtimes; Mentor engineers and develop technical leaders; Coordinate end-to-end optimization of AI models and inference runtimes; Establish team processes for system-level debugging and performance optimization.
Seniority
Senior, hands-on IC with leadership responsibilities