CareerPlanSign in

Software Engineer, Inference (AI Data Engineering)

Palo Alto - 1530💼 Full-time💰 $1–$1🗓 2026-08-19 → 2026-09-26

Core

Design and optimize large-scale, high-performance AI inference platforms to serve mission-critical models for SpaceX's launch vehicles and Starlink.

Role type

Senior IC AI Data Engineer (Inference Systems)

Builds

Distributed model serving infrastructure, inference engines, and high-concurrency serving systems for internal SpaceX applications.

Domain

Aerospace / AI Infrastructure / High-Performance Computing

Deliverable

production ML models

Required skills

Rust or C++, distributed systems design, full stack or backend development, low-level systems programming, GPU kernel optimization, model inference acceleration (quantization, speculative decoding), CI/CD infrastructure, observability, database management (PostgreSQL, ClickHouse, MongoDB), gRPC, containerization (Docker, Kubernetes), Python or Go.

Preferred skills

LLM inference engines (SGLang, vLLM, Triton, TensorRT-LLM), agent SDKs and orchestration frameworks, profiling and performance tuning.

Responsibilities

Develop reliable, high-throughput inference systems; architect scalable distributed infrastructure for model serving; optimize latency and throughput under production workloads; build high-concurrency serving systems with 100% uptime; own end-to-end components like request routing and SDK development; benchmark and accelerate inference engines; develop custom tracing and debugging tools; create robust CI/CD infrastructure; collaborate to integrate inference capabilities into broader workflows.

Seniority

Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.