Senior Principal AI Infrastructure Architect
Core
Designing complex AI platform and managed-service solutions, focusing on hardware foundations for enterprise-scale training, fine-tuning, and inference workloads.
Role type
Senior Principal AI Infrastructure Architect
Builds
Enterprise-scale AI infrastructure solutions including GPU clusters, storage tiers, and AI fabrics for sovereign AI and AI Factory clients.
Domain
AI Infrastructure, Datacenter Engineering, High-Performance Computing
Deliverable
production ML models
Required skills
AI hardware architecture (GPU/accelerators, CPUs), AI-class storage systems, AI networking (InfiniBand, RoCE, NVLink), AI software stack orchestration, datacenter facilities engineering, business case modeling, presales consulting
Preferred skills
Vendor certifications (NVIDIA, Dell, Cisco, etc.), Scaled Agile certification, cloud/hybrid deployment patterns
Technologies
NVIDIA H100/H200/B200, AMD Instinct, Intel Gaudi 3, NVIDIA DGX/HGX, InfiniBand NDR/XDR, NVMe, Kubernetes, Slurm, Kubeflow, CUDA, ROCm
Responsibilities
Lead end-to-end design of large AI infrastructure solutions, architect reference designs, size and validate GPU clusters, define datacenter design, drive presales process, translate client ambitions into hardware roadmaps, lead integration of compute/storage/networking/software, build business cases and TCO models, define architectural principles, develop transition roadmaps.
Seniority
Senior Principal, hands-on IC with strategic vision