Lead AI Engineer (LLM Inference)
Core
Design, develop, and optimize large-scale LLM inference systems and AI services to transform banking workflows and customer interactions.
Role type
Senior IC Lead AI Engineer (LLM Inference)
Builds
Production AI services including foundation model training, LLM inference, similarity search, guardrails, and model evaluation.
Domain
Financial services / Large Language Model Inference / Cloud Infrastructure
Deliverable
production ML models
Required skills
LLM Inference, Similarity Search, VectorDBs, Guardrails, Model Optimization, Scalable Cloud Deployment, Python, C++, C#, Java, Golang, PyTorch, AWS, Azure
Preferred skills
Hardware utilization optimization, Advanced inference techniques, AI research application
Technologies
AWS Ultraclusters, Huggingface, Nemo Guardrails, PyTorch, VectorDBs
Responsibilities
Design and deploy scalable AI solutions on cloud platforms; Optimize training and inference software for latency, throughput, and cost; Collaborate with cross-functional teams to deliver AI-driven products; Shape the technical vision and roadmap for foundational AI systems.
Seniority
Senior, hands-on IC with technical leadership