CoDesign & NextGen Performance Engineer
Core
Characterizing, analyzing, and optimizing the performance of state-of-the-art AI models running on Cerebras' breakthrough Wafer Scale Engine (WSE) hardware.
Role type
Senior IC performance engineer (AI hardware/software stack)
Builds
Production ML inference systems on Cerebras WSE
Domain
AI hardware / Computer Architecture / High-Performance Computing
Deliverable
production ML models
Required skills
Computer architecture, Kernel optimization, Low-level deep learning math, Performance profiling, C++, Python
Preferred skills
CPU/GPU simulator experience, Cluster runtime debugging
Technologies
Cerebras WSE, C++, Python
Responsibilities
Bring up and optimize performance on new generations of the Cerebras WSE; Build performance models (kernel-level, end-to-end) to estimate performance of ML models; Optimize and debug kernel micro code and compiler algorithms; Debug and understand runtime performance on the system and cluster; Develop tools and infrastructure to visualize performance data.
Seniority
Senior, hands-on IC