AI Engineer 5 (LLM Gateway, FM Hosting)
Core
Design, develop, test, deploy, and support large-scale AI software components including foundation model training, LLM inference, agents, multi-agent workflows, similarity search, guardrails, and observability for banking applications.
Role type
Senior IC AI Engineer (LLM Gateway & Foundation Model Hosting)
Builds
Scalable, high-performance AI infrastructure and multi-model orchestration pipelines integrating LLMs, vector search, and domain-specific models.
Domain
Financial Services / Large Language Models / Cloud Infrastructure
Deliverable
production ML models
Required skills
Foundation model optimization, LLM inference engineering, multi-agent workflow design, vector search implementation, model evaluation and governance, cost-performance analysis, GPU utilization management, technical leadership, system architecture
Preferred skills
Agentic AI system development, heterogeneous AI system integration, ethical AI standards enforcement, dynamic inference strategy, model compression, hardware right-sizing
Technologies
AWS Ultraclusters, Huggingface, VectorDBs, PyTorch, CUDA, Python, Go, Scala, Java, C++, C#
Responsibilities
Design and implement multi-model orchestration pipelines; establish cost-performance governance reviews; lead team design councils; mentor Principal and Manager-level engineers; optimize foundation model performance for scalability and latency.
Seniority
Senior, hands-on IC with mentorship responsibilities