Engineering Manager, AI Platform - Managed AI
Core
Lead and scale a team of engineers building a next-generation platform for the full lifecycle of Large Language Models (LLMs), focusing on scalable, fault-tolerant infrastructure.
Role type
Engineering Manager, AI Platform
Builds
Next-generation AI platform for LLM lifecycle (task queues, model management, scheduling)
Domain
AI Infrastructure / Cloud Services
Deliverable
production ML models
Required skills
Team leadership, distributed systems, cloud-native environments, container orchestration, SOA, hiring, technical roadmap execution
Preferred skills
CPU & GPU performance, inference frameworks, LLM systems, Python/GoLang/Rust, Kubernetes, gRPC, observability stacks, open-source AI ecosystems (vLLM, Hugging Face, Triton)
Technologies
Kubernetes, gRPC, vLLM, Hugging Face, Triton, Python, GoLang, Rust
Responsibilities
Lead and mentor high-caliber software engineers; Define and execute the AI roadmap; Oversee architecture of core AI services; Ensure delivery of scalable systems; Work cross-functionally with Product, Infrastructure, and GTM stakeholders
Seniority
Manager, hands-on leadership