Neuron Runtime Software Engineer
Core
Develop and maintain high-performance runtime libraries, drivers, and profilers for machine learning applications on AWS Inferentia and Trainium hardware.
Role type
Senior IC software engineer (ML runtime & hardware stack)
Builds
Neuron Runtime, ML kernels, and profiling tools for Trainium/Inferentia accelerators
Domain
Cloud-scale machine learning infrastructure and hardware acceleration (via careerplan.io/jobs/g2bb815c-neuron-runtime-software-engineer-at-annapurna-labs-u-s-inc)
Required skills
C++ programming, distributed systems architecture, system reliability design, AWS services (EC2, ECS, Lambda, CloudWatch, S3), full software development lifecycle management, performance optimization
Preferred skills
ML framework integration (PyTorch, JAX, XLA), compiler optimization, cross-functional product leadership
Technologies
C++, AWS Lambda, CloudWatch, EC2, PyTorch, JAX, XLA, DevOps
Responsibilities
Design and deploy Neuron Runtime components, enhance ML kernel and framework performance, manage full development lifecycle for scalability and reliability, collaborate on C++ compiler improvements for hardware tuning, drive profiler support across multiple frameworks, partner with leadership on product direction
Seniority
Senior, hands-on IC with architectural responsibilities
