CareerPlanGet AI match score →

Principal AI/ML System Software Engineer

Santa Clara💼 Full-time🗓 2026-05-15 → 2026-07-31

Core

Building and scaling the next-generation AI deployment software stack and deployment infrastructure for an AI compute engine.

Role type

Principal System Software Engineer (AI Inference Execution)

Builds

Production AI inference software and deployment infrastructure

Domain

AI compute hardware/software co-design and inference execution

Deliverable

production ML models

Required skills

C/C++/Python development, Linux environment, distributed high-performance software design, computer architecture, data structures, machine learning fundamentals

Preferred skills

Inference servers/model serving frameworks, deep learning frameworks, deep learning runtimes, distributed systems collectives, software testing fundamentals, MLOps tools, Kubernetes, Ray, startup/incubation experience, cloud provider or AI compute company experience

Technologies

TensorRT-LLM, vLLM, SGLang, PyTorch, TensorFlow, ONNX Runtime, TensorRT, NCCL, OpenMPI, Kubernetes, Ray

Responsibilities

Develop, enhance, and maintain next-generation AI deployment software; work with system software experts to build deployment infrastructure; collaborate with ML, compilers, and hardware experts; build and scale software deliverables within tight development windows

Seniority

Principal, hands-on IC with leadership

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗