Senior Machine Learning Applications and Compiler Engineer, LPX
Core
Develop algorithms and optimizations for NVIDIA's LPX inference and compiler stack, mapping neural network workloads to future platforms.
Role type
Senior IC machine learning applications and compiler engineer
Builds
High-performance runtime and compiler components for end-to-end inference optimization on NVIDIA systems
Domain
Computer systems, compilers, deep learning, hardware-software codesign
Deliverable
production ML models
Required skills
systems-level programming, compiler/runtime development, IR design, optimization passes, code generation, LLVM/MLIR, deep learning frameworks, parallel/heterogeneous compute architectures, performance profiling
Preferred skills
MLIR-based compilers, spatial/dataflow architectures, static scheduling, pipeline/tensor parallelism, open-source contributions, research publications
Technologies
C/C++, Rust, LLVM, MLIR, TensorFlow, PyTorch, ONNX
Responsibilities
build and maintain runtime/compiler components, define workload mappings, extend NVIDIA SW ecosystem, benchmark/profile performance, collaborate with hardware architects, prototype compilation techniques
Seniority
Senior, hands-on IC