Deep Learning Solution Architect
Core
Design, optimize, and deliver production-grade generative AI solutions for enterprise customers using NVIDIA's hardware and software ecosystem.
Role type
Senior Solution Architect (Generative AI / LLMs)
Builds
End-to-end LLM solutions including pretraining, fine-tuning, high-performance inference, RAG workflows, and agentic inference orchestration.
Domain
Enterprise Generative AI, Large Language Models, GPU Computing
Deliverable
production ML models
Required skills
LLM pretraining and fine-tuning, distributed optimization, RAG workflow design, agentic inference orchestration, GPU cluster architecture, PyTorch, Hugging Face Transformers
Preferred skills
NVIDIA TRT-LLM, Megatron-LM, NVIDIA NeMo, LLM quantization, KV Cache tuning, Docker, Kubernetes, multi-GPU parallelism
Responsibilities
Architect end-to-end LLM solutions, collaborate with customers on business challenges, lead LLM training and performance tuning, design and integrate RAG and agentic pipelines, support pre-sales technical activities
Seniority
Senior, hands-on IC