Senior AI Engineer (FM Hosting, LLM Inference)
Core
Design, develop, test, deploy, and maintain AI software components including foundation model training, large language model inference, similarity search, and model evaluation to transform associate workflows and customer interactions.
Role type
Senior IC AI Engineer (LLM Inference & Optimization)
Builds
Scalable AI systems, foundation models, and LLM inference pipelines for banking operations and customer experiences.
Domain
Financial Services / Large Language Models / Cloud Infrastructure
Deliverable
production ML models
Required skills
LLM inference optimization, similarity search, model evaluation, cloud platform deployment, Python, Go, Scala, Java, C++, C#, PyTorch, AWS, Azure, hardware efficiency tuning, cross-functional leadership
Preferred skills
AI research, novel technique application, complex system design
Technologies
AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, AWS, Azure, C#, Golang, Java, LLM, Machine Learning, Python, Scala
Responsibilities
Collaborate with engineers, scientists, and product managers to deliver AI-driven products; Design and implement advanced LLM optimization techniques for scalability and cost-efficiency; Contribute to the technical vision and roadmap of foundational AI systems.
Seniority
Senior, hands-on IC with leadership responsibilities