Senior Performance Engineer, Efficiency Red Team
Core
Hunting for hidden waste across the world's largest compute infrastructure to turn findings into freed capacity and cost savings.
Role type
Senior IC performance engineer (efficiency/red team)
Builds
Fleet-wide optimization playbooks, automated profiling tools, and efficiency frameworks for Amazon's compute stack.
Domain
Cloud infrastructure, hyper-scale systems, Generative AI, hardware/software intersection
Deliverable
production ML models | infrastructure
Required skills
systems programming, performance profiling, root cause analysis, capacity planning, GPU memory management, language runtime internals, C/C++/Rust/Java
Preferred skills
GPU profiling (CUDA), generative AI infrastructure, language runtime internals (JVM), open source contributions
Technologies
CUDA, JVM, C, C++, Rust, Java
Responsibilities
Apply first-principles analysis across CPU, DRAM, GPU, storage, I/O, and networking layers; Design and lead rigorous investigations to establish root causes; Build proof-of-concept implementations for novel efficiency approaches; Develop repeatable patterns and practices for fleet-wide optimization; Quantify business impact at Amazon scale; Collaborate across organizational boundaries; Contribute findings to internal profiling tools and automated optimization agents.
Seniority
Senior, hands-on IC with mentorship/tech lead experience