Senior System Architect, Enterprise Reference Architectures
Core
Define, design, and validate enterprise AI factory reference architectures for NVIDIA accelerated computing, networking, storage, and AI software.
Role type
Senior System Architect (Enterprise Reference Architectures)
Builds
On-prem cloud-native platforms, cluster reference builds, and scalable AI/ML systems for enterprise and partner deployment.
Domain
Enterprise AI infrastructure, Data Center Architecture, High-Performance Computing (HPC)
Deliverable
production ML models | product features | infrastructure
Required skills
Data center architecture, Microservices, Distributed systems design, GPU-accelerated systems, High-performance networking, Kubernetes, Event-driven architectures, Cloud-native systems, API strategies, Middleware integrations
Preferred skills
Product strategy, Customer requirements translation, Technical mentorship, Trade-off analysis (performance, scalability, TCO)
Technologies
NVIDIA AI software, InfiniBand, RDMA, RoCE, Kafka, RabbitMQ, Docker
Responsibilities
Define full-stack enterprise AI factory baseline architectures; Architect reference builds for end-to-end software systems; Develop and validate scalable cluster designs for training, fine-tuning, inference, and HPC workloads; Work with Product Management and Engineering to develop product direction and roadmap inputs; Evaluate tradeoffs across performance, security, power, cooling, and cost; Build comprehensive API strategies and orchestration workflows.
Seniority
Senior, hands-on IC with strategy & mentorship