Compute Server Platform Architect
Core
Design and own the server-side platform architecture for Cerebras CS3-based AI clusters to ensure predictable performance, scalability, and reliability for training and inference workloads.
Role type
Senior Compute Server Platform Architect
Builds
x86 server fleets and cluster configurations that co-design with Cerebras accelerators for high-speed AI inference and training
Domain
AI/ML infrastructure, high-performance computing (HPC), large-scale distributed systems
Deliverable
production ML models | infrastructure
Required skills
x86 server architecture, Linux systems performance tuning, high-performance IO (RDMA/RoCE/NVMe), capacity modeling, vendor engagement, cross-functional technical leadership
Preferred skills
C/C++/Python, emerging server technologies (CXL, SmartNIC/DPU), OS/BIOS/firmware baseline definition
Technologies
x86, Linux, RDMA, RoCE, NVMe, CXL, AMD, Intel, ARM
Responsibilities
Define server types, configurations, and lifecycle strategy; specify platform configurations (CPU, memory, PCIe, NIC); translate software flows into hardware requirements; develop and validate performance/scaling models; define OS/BIOS/firmware baselines; lead technical vendor engagements; define qualification criteria and execute qualification plans.
Seniority
Senior, hands-on IC with technical leadership