1,062 open positions
Senior Performance Co-design Engineer, LLM Training↗
GoogleSunnyvale, CA, US
$174k–$174k1mo
Lead hardware/software co-design and performance modeling for LLM training workloads to optimize Google's custom AI silicon.
Software Engineering Manager, ML Kernel Performance, AWS Neuron, Annapurna Labs↗
AmazonToronto, ON, CA
$171k–$286k1mo
Architect and implement business-critical features for AWS Neuron SDK to accelerate deep learning and GenAI workloads on Inferentia and Trainium hardware.
Lead Software Engineer, ML Network Stack - Annapurna Labs
AmazonCupertino, California, United States
$193k–$193k1mo
Lead the design and architecture of the network stack for EC2 distributed AI/ML systems, enabling NVIDIA GPUs to work with AWS EC2 machines.
Software Development Engineer, ML Systems, Annapurna Labs
AmazonNew York, New York, United States
$158k–$214k1mo
Building AI agents, tools, and models to simplify and accelerate customer adoption of AWS Neuron (software stack for Trainium/Inferentia ML chips) and optimizing ML workloads on Neuron.
Senior Performance Co-Design Engineer, LLM Serving↗
GoogleSunnyvale, CA, US
$174k–$174k1mo
Conduct LLM serving performance studies and optimize inference latency/throughput on custom hardware (TPUs) by partnering with hardware architects.
Senior Software Engineer, Fleet-level ML Performance↗
GoogleSunnyvale, CA, US
$174k–$174k2mo
Conduct end-to-end performance analysis of critical ML workloads (e.g., Gemini) on next-generation TPU systems using advanced simulation to evaluate hardware/software trade-offs and guide chip architecture.
Software Development Engineer AI/ML, Inference Model Enablement, AWS Neuron
AmazonCupertino, California, United States
$193k–$262k2mo
Onboard and optimize state-of-the-art open-source and customer LLMs (dense and MoE) for inference on Trainium accelerators, while advancing inference usability and quality through features, infrastructure optimization, tools, and automation.
← Select a job to preview