Staff AI Systems Engineer
Core
Architect, deploy, and manage critical infrastructure services for large-scale AI model training and inference, bridging raw data to high-performance models for an all-electric aircraft company.
Role type
Staff AI Systems Engineer (ML Infrastructure)
Builds
Resilient distributed AI training infrastructure, low-latency inference platforms, and end-to-end MLOps tooling.
Domain
Aerospace / AI Infrastructure / High-Performance Computing
Deliverable
infrastructure
Required skills
Distributed AI model training, MLOps (MLflow), Multi-cloud compute management, Kubernetes, LLM serving optimization, Data pipelines for AI
Preferred skills
Audio processing, Speech-to-text, Automatic Speech Recognition (ASR), Aerospace/Aviation domain knowledge
Technologies
AWS, Nebius AI Cloud, Docker, Kubernetes, SkyPilot, vLLM, SGLang, SQL, NoSQL, Parquet
Responsibilities
Deploy and scale resilient infrastructure for distributed AI training; Maintain end-to-end MLOps tooling; Optimize compute and inference on multi-cloud GPU environments; Partner with researchers to productionize models and debug hardware-software bottlenecks.
Seniority
Staff, hands-on IC with strategic scope