CareerPlanGet AI match score →

Senior AI Engineer (FM Hosting, LLM Inference)

USA💼 Full-time💰 $229,900–$229,900🗓 2026-07-14 → 2026-07-30

Core

Design, develop, test, deploy, and maintain AI software components including foundation model training, large language model inference, similarity search, and model evaluation to transform associate workflows and customer interactions.

Role type

Senior IC AI Engineer (LLM Inference & Optimization)

Builds

Scalable AI systems, foundation models, and LLM inference pipelines for banking operations and customer experiences.

Domain

Financial Services / Large Language Models / Cloud Infrastructure

Deliverable

production ML models

Required skills

LLM inference optimization, similarity search, model evaluation, cloud platform deployment, Python, Go, Scala, Java, C++, C#, PyTorch, AWS, Azure, hardware efficiency tuning, cross-functional leadership

Preferred skills

AI research, novel technique application, complex system design

Technologies

AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, AWS, Azure, C#, Golang, Java, LLM, Machine Learning, Python, Scala

Responsibilities

Collaborate with engineers, scientists, and product managers to deliver AI-driven products; Design and implement advanced LLM optimization techniques for scalability and cost-efficiency; Contribute to the technical vision and roadmap of foundational AI systems.

Seniority

Senior, hands-on IC with leadership responsibilities

Sourced via devitjobs · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Devitjobs ↗