Lead AI Engineer (FM Hosting, LLM Inference)
Core
Design, develop, test, deploy, and support AI software components including foundation model training, large language model inference, similarity search, guardrails, model evaluation, experimentation, governance, and observability to build scalable, high-performance AI infrastructure.
Role type
Lead AI Engineer (LLM Inference & Foundation Models)
Builds
Proprietary AI solutions and platforms that empower teams across the company to enhance products with transformative AI capabilities.
Domain
Banking / Financial Services / Large Language Models / Cloud Infrastructure
Deliverable
production ML models
Required skills
LLM inference optimization, foundation model training, similarity search, vector databases, model guardrails, model evaluation, experimentation, governance, observability, system design, Python, Go, Scala, Java
Preferred skills
Cloud platform deployment (AWS, GCP, Azure), scalable AI services delivery, hardware utilization optimization, latency and throughput improvement, cost reduction in AI systems
Technologies
AWS Ultraclusters, Huggingface, VectorDBs, Nemo Guardrails, PyTorch
Responsibilities
Partner with cross-functional teams to deliver AI-powered products; Invent and introduce state-of-the-art LLM optimization techniques; Contribute to the technical vision and long-term roadmap of foundational AI systems.
Seniority
Lead, hands-on IC with strategic roadmap contribution