AI Engineer, FM Hosting, LLM Inference
Core
Design, develop, test, deploy, and support large-scale AI systems including LLM inference, agentic workflows, and multi-model orchestration to improve associate and customer experiences.
Role type
Senior IC AI Engineer (LLM Inference & Agentic Systems)
Builds
Production AI systems, multi-model orchestration pipelines, and agentic workflows for banking customers.
Domain
Financial Services / Large Language Models / AI Infrastructure
Deliverable
production ML models
Required skills
LLM inference optimization, agentic AI system design, multi-model orchestration, cost-performance governance, model compression, hardware utilization optimization, vector search, guardrails, system architecture, technical leadership, mentorship
Preferred skills
Cloud platform deployment (AWS/GCP/Azure), foundation model training, ethical AI standards enforcement, dynamic inference strategies, right-sizing models and instances
Technologies
PyTorch, CUDA, Golang, Java, Python, C++, C#, AWS Ultraclusters, Hugging Face, VectorDBs
Responsibilities
Partner with product and engineering teams to deliver AI-powered products; design and optimize multi-model orchestration pipelines; lead cost-performance governance reviews; mentor Principal- and Manager-level engineers; establish technical consistency and compliance with AI engineering standards.
Seniority
Senior, hands-on IC with leadership responsibilities

