CareerPlanGet AI match score →

CoDesign & NextGen Performance Engineer

Headquarters/Sunnyvale Office💼 Full-time🗓 2026-07-14 → 2026-07-31

Core

Characterizing, analyzing, and optimizing the performance of state-of-the-art AI models running on Cerebras' breakthrough Wafer Scale Engine (WSE) hardware.

Role type

Senior IC performance engineer (AI hardware/software stack)

Builds

Production ML inference systems on Cerebras WSE

Domain

AI hardware / Computer Architecture / High-Performance Computing

Deliverable

production ML models

Required skills

Computer architecture, Kernel optimization, Low-level deep learning math, Performance profiling, C++, Python

Preferred skills

CPU/GPU simulator experience, Cluster runtime debugging

Technologies

Cerebras WSE, C++, Python

Responsibilities

Bring up and optimize performance on new generations of the Cerebras WSE; Build performance models (kernel-level, end-to-end) to estimate performance of ML models; Optimize and debug kernel micro code and compiler algorithms; Debug and understand runtime performance on the system and cluster; Develop tools and infrastructure to visualize performance data.

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗