AI Inference Engineer
Core
Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform, translating business goals into reliable, observable services.
Role type
Forward Deployed Engineer (hands-on IC with product/solution engineering aspects)
Builds
Production-ready model servers, custom workflows, and end-to-end AI/ML solutions for customers
Domain
AI/ML Infrastructure & Customer Solutions
Deliverable
production ML models | product features
Required skills
Python, general-purpose programming languages, AI/ML pipelines, model deployment, problem framing, evaluation, monitoring, project management
Preferred skills
Building or optimizing AI/ML projects
Technologies
Python, Docker, ComfyUI
Responsibilities
Design, implement, and deploy Baseten solutions end-to-end; develop and maintain software systems and product features; turn vague objectives into clear specs and PoCs; optimize and enhance AI/ML projects; own products and customer projects end-to-end
Seniority
Mid-level, hands-on IC