CareerPlanSign in

AI Inference Engineer

San Francisco💼 Full-time🗓 2026-08-03 → 2026-09-26

Core

Partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten's platform, translating business goals into reliable, observable services.

Role type

Forward Deployed Engineer (hands-on IC with product/solution engineering aspects)

Builds

Production-ready model servers, custom workflows, and end-to-end AI/ML solutions for customers

Domain

AI/ML Infrastructure & Customer Solutions

Deliverable

production ML models | product features

Required skills

Python, general-purpose programming languages, AI/ML pipelines, model deployment, problem framing, evaluation, monitoring, project management

Preferred skills

Building or optimizing AI/ML projects

Technologies

Python, Docker, ComfyUI

Responsibilities

Design, implement, and deploy Baseten solutions end-to-end; develop and maintain software systems and product features; turn vague objectives into clear specs and PoCs; optimize and enhance AI/ML projects; own products and customer projects end-to-end

Seniority

Mid-level, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.