Senior Machine Learning Engineer (Large Systems)
Core
Developing and optimizing AI models tailored to specialized hardware for large-scale systems where performance is critical.
Role type
Senior IC machine learning engineer (large systems)
Builds
Reference applications, optimized software kernels, and next-generation AI hardware
Domain
Semiconductor hardware, AI compute stack, distributed training and inference
Deliverable
production ML models
Required skills
Deep learning frameworks (PyTorch/JAX), Python/C++ development, model training to optimization, experimental design and execution, performance bottleneck analysis
Preferred skills
MLOps for Kubernetes, LLM production systems, low-precision arithmetic, C++/Triton/CUDA kernel development, distributed training across 64+ accelerators, HPC systems (Infiniband/NVLink/RoCE), open-source contributions
Technologies
PyTorch, JAX, Python, C++, Triton, CUDA, Kubernetes, Infiniband, NVLink, RoCE
Responsibilities
Implement and optimize ML models for 1000s of accelerators, test and evaluate software releases, benchmark models to identify bottlenecks, design and execute novel AI experiments, collaborate on next-gen AI hardware, engage with the AI community
Seniority
Senior, hands-on IC