CareerPlanGet AI match score →

Lead AI Engineer (FM Hosting, LLM Inference)

3 Locations💼 Full-time💰 $197,300–$197,300🗓 2026-05-12 → 2026-07-31

Core

Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability to build scalable, high-performance AI infrastructure.

Role type

Lead AI Engineer (LLM Inference & Foundation Models)

Builds

Proprietary AI solutions and platforms that empower teams across the company to enhance products with transformative AI capabilities.

Domain

Banking / Financial Services / Large Language Models / Cloud Infrastructure

Deliverable

production ML models

Required skills

LLM inference optimization, foundation model training, similarity search, vector databases, model guardrails, model evaluation, experimentation, governance, observability, system design, Python, Go, Scala, Java

Preferred skills

Cloud platform deployment (AWS, GCP, Azure), scalable AI services delivery, hardware utilization optimization, latency and throughput improvement, cost reduction in AI systems

Technologies

AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch

Responsibilities

Partner with cross-functional teams to deliver AI-powered products; Invent and introduce state-of-the-art LLM optimization techniques; Contribute to the technical vision and long-term roadmap of foundational AI systems.

Seniority

Lead, hands-on IC with strategic roadmap contribution

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Workday ↗