Senior Technical Program Manager – AI Infrastructure, Site Operations
Core
Own end-to-end technical programs for data center and site operations supporting Cerebras' AI Cloud and customer deployments.
Role type
Senior Technical Program Manager (AI Infrastructure & Site Operations)
Builds
Reliable deployment, operation, and scaling of Cerebras Wafer-Scale Engine systems in data centers.
Domain
AI Hardware Infrastructure / Data Center Operations
Deliverable
production ML models | infrastructure
Required skills
Technical Program Management, cross-functional infrastructure program leadership, data center power and cooling fundamentals, network and storage basics, hardware-centric platforms, operational metrics definition, executive-level communication
Preferred skills
AI/ML or HPC infrastructure, accelerator-based infrastructure, high-density or liquid-cooled data centers, colocation provider management, incident management and reliability
Technologies
Cerebras Wafer-Scale Engine, AI Cloud Infrastructure, Network & Storage Engineering, Facilities & Power Systems
Responsibilities
Own end-to-end technical programs for data center and site operations; Act as single-threaded owner across Hardware & Systems Engineering, AI Cloud Infrastructure & Operations, Network & Storage Engineering, and Facilities; Drive site readiness for Cerebras Wafer-Scale Engine systems; Partner on installation, commissioning, change management, and break/fix workflows; Lead incident reviews and postmortems; Define and own operational metrics and KPIs; Build executive-level dashboards and reporting; Establish program governance, risk tracking, and RACI clarity; Present program status, metrics, and operational risks to senior leadership
Seniority
Senior, hands-on IC