Senior Solutions Architect, Generative AI Deployment and AIOps
Core
Trusted technical advisor helping customers build and deploy Generative AI and LLM inference solutions using NVIDIA hardware and software platforms.
Role type
Senior Solutions Architect (Generative AI Deployment and AIOps)
Builds
Inference workloads for Generative AI and Large Language Models (LLMs) on NVIDIA computing platforms
Domain
AI Infrastructure / Generative AI / High-Performance Computing
Deliverable
production ML models
Required skills
Deep Learning frameworks (PyTorch, TensorFlow), Python programming, GPU orchestration, Kubernetes management, MLOps, LLM inference theory and practice, containerization, observability
Preferred skills
DL training at scale, NVIDIA NIM/Dynamo/TensorRT/TensorRT-LLM, C/C++ programming, parallel programming, distributed computing
Technologies
PyTorch, TensorFlow, Python, Kubernetes, NVIDIA GPUs, NVIDIA NIM, NVIDIA Dynamo, TensorRT, TensorRT-LLM, C/C++
Responsibilities
Partner with engineering and product teams to define high-value AI solutions; Engage with developers and researchers to understand technical needs; Collaborate with lighthouse customers and partners to adopt NVIDIA technology; Analyze performance and power efficiency of AI inference workloads on Kubernetes
Seniority
Senior, hands-on IC