Software Engineer, Inference (AI Data Engineering)
Core
Design and optimize large-scale, high-performance AI inference platforms to serve mission-critical models for SpaceX's launch vehicles and Starlink.
Role type
Senior IC AI Data Engineer (Inference Systems)
Builds
Distributed model serving infrastructure, inference engines, and high-concurrency serving systems for internal SpaceX applications.
Domain
Aerospace / AI Infrastructure / High-Performance Computing
Deliverable
production ML models
Required skills
Rust or C++, distributed systems design, full stack or backend development, low-level systems programming, GPU kernel optimization, model inference acceleration (quantization, speculative decoding), CI/CD infrastructure, observability, database management (PostgreSQL, ClickHouse, MongoDB), gRPC, containerization (Docker, Kubernetes), Python or Go.
Preferred skills
LLM inference engines (SGLang, vLLM, Triton, TensorRT-LLM), agent SDKs and orchestration frameworks, profiling and performance tuning.
Responsibilities
Develop reliable, high-throughput inference systems; architect scalable distributed infrastructure for model serving; optimize latency and throughput under production workloads; build high-concurrency serving systems with 100% uptime; own end-to-end components like request routing and SDK development; benchmark and accelerate inference engines; develop custom tracing and debugging tools; create robust CI/CD infrastructure; collaborate to integrate inference capabilities into broader workflows.
Seniority
Senior, hands-on IC