Senior AI Engineer
Core
Building a centralized AI function to own the model layer, including fine-tuning, distilling, and operating inference infrastructure for DV-specific tasks.
Role type
Senior IC machine-learning engineer (LLM infrastructure)
Builds
Model gateway routing inference across open and closed models; on-prem inference infrastructure; distillation pipelines for task-specific open models
Domain
Financial services / Proprietary trading / AI infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, production fine-tuning of open-weight models, LLM serving on-prem, GPU infrastructure management, Kubernetes, model evaluation and regression testing
Preferred skills
Quantization, PEFT/LoRA, model gateway design, financial services experience, open model ecosystem familiarity
Technologies
vLLM, TGI, Triton, Kubernetes, Llama, Qwen, Mistral
Responsibilities
Build and operate a model gateway routing inference across open and closed models with cost, latency, and quality tracking; Design and run distillation pipelines using frontier model outputs to generate training data; Fine-tune and evaluate open-weight models for DV-specific tasks; Deploy and maintain on-prem inference infrastructure on Kubernetes; Build model evaluation frameworks for quality, cost, latency, and regression; Define criteria and tooling for model selection
Seniority
Senior, hands-on IC