Engineering Manager, Foundation Model Inference (FMAPI)
Core
Lead infrastructure engineering teams building scalable systems for serving and optimizing large-scale foundation model inference workloads (partner and self-hosted models) for enterprise customers.
Role type
Senior Engineering Manager, AI Infrastructure
Builds
Real-time, provisioned throughput, and batch inference platforms for LLMs
Domain
Generative AI, Large Language Models (LLMs), Distributed Systems
Deliverable
production ML models
Required skills
Team leadership, distributed systems architecture, AI/ML infrastructure, large-scale backend services, roadmap planning, service health management, incident response, hiring
Preferred skills
None stated
Technologies
None explicitly listed
Responsibilities
Lead and grow a team of infrastructure engineers; Shape product roadmap for inference use cases; Partner with product and engineering leadership; Build an inclusive, high-performing team; Maintain technical bar for architecture and operational excellence
Seniority
Senior, hands-on IC leader