Senior AI Platform Engineer
Core
Define and deliver infrastructure strategy for IQVIA's Large Language Model (LLM) programmes, transforming research innovations into secure, scalable, production-ready AI solutions.
Role type
Senior IC AI Platform Engineer (LLM Infrastructure)
Builds
High-performance compute environments, model lifecycle platforms, and knowledge graph infrastructure for enterprise AI systems.
Domain
Healthcare + Large Language Models (LLM) + Distributed Computing
Deliverable
production ML models
Required skills
LLM architecture design, GPU infrastructure management, distributed training strategies, model optimization (quantization, mixed precision), AWS cloud services, workload orchestration (Slurm, Kubernetes, Ray), high-performance computing, infrastructure automation, cross-functional technical leadership.
Preferred skills
Vendor selection for AI platforms, mentoring engineers, bridging research and production environments.
Technologies
CUDA, cuDNN, NCCL, PyTorch, vLLM, TensorRT-LLM, NVIDIA NIM, SGLang, NVIDIA Nsight, DCGM, Slurm, Kubernetes, Ray, AWS.
Responsibilities
Own the AI platform and infrastructure roadmap; design and deliver high-performance compute environments; optimize LLM training and inference workloads; establish model and data lifecycle capabilities; lead knowledge graph infrastructure evolution; serve as primary technical coordination point across teams; provide technical leadership for vendor selection and technology partnerships.
Seniority
Senior, hands-on IC with strategic leadership