AI Engineer: LLM & Fine-Tuning
Core
Design, develop, test, deploy, and support large-scale production AI systems including LLM inference, agentic workflows, and multi-model orchestration to improve banking products.
Role type
Senior IC AI Engineer (LLM & Fine-Tuning)
Builds
Scalable, high-performance AI systems for banking associates and customers
Domain
Financial Services / Large Language Models / Cloud Infrastructure
Deliverable
production ML models
Required skills
LLM inference, agentic AI systems, multi-agent workflows, model training, vector databases, model optimization, cost-performance governance, technical leadership, mentorship
Preferred skills
Leading AI system development, tradeoff decisions (cost/latency/throughput/accuracy), deploying scalable AI on cloud platforms, building AI/ML capabilities (guardrails, memory), optimizing training/inference software, architecting heterogeneous AI systems, ethical AI deployment standards, dynamic inference strategies, model compression, right-sizing models
Technologies
PyTorch, AWS Ultraclusters, Hugging Face, CUDA, Python, C++, C#, Java, Golang, Scala
Responsibilities
Design and implement multi-model orchestration pipelines combining LLMs, vector search, and domain-specific models; lead design councils to maintain technical consistency; mentor Principal- and Manager-level AI engineers; establish and lead cost-performance governance reviews tracking GPU utilization and inference cost efficiency; invent advanced foundation model optimization techniques for scalability and latency; partner with product managers to deliver AI-powered products.
Seniority
Senior, hands-on IC with leadership and mentorship responsibilities
