CareerPlanGet AI match score →

Performance & Systems Engineer, Codex

San Francisco💼 Full-time🗓 2026-05-01 → 2026-07-31

Core

Optimizing the performance, latency, and cost of a complex AI system stack for code generation and agentic reasoning products.

Role type

Senior IC performance & systems engineer (LLM inference & cloud infrastructure)

Builds

High-leverage changes across infrastructure, modeling, and product layers to improve system speed and cost efficiency.

Domain

AI/ML systems, cloud infrastructure, agentic work management

Deliverable

production ML models | infrastructure

Required skills

LLM inference optimization, cloud orchestration, system profiling, container orchestration, latency reduction, cost optimization, cross-stack debugging

Preferred skills

Ambiguity navigation, holistic performance thinking, tooling development

Technologies

LLM inference engines, cloud orchestration platforms, container orchestration tools

Responsibilities

Identify and resolve inefficiencies across the system stack from agent behavior to inference to orchestration; Build tooling to measure, profile, and optimize system performance at scale; Collaborate with researchers and engineers to implement high-ROI changes improving latency and cost.

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗