Member of Engineering (Compute)
Core
Design and develop internal GPU workload scheduling systems and inference serving stacks to optimize utilization and stability for AI research.
Role type
Senior IC systems engineer (GPU scheduling & inference)
Builds
Internal scheduling systems, inference control planes, and GPU lifecycle management tooling
Domain
AI infrastructure, distributed systems, GPU compute
Deliverable
production ML models | infrastructure
Required skills
Go, distributed systems, Kubernetes internals, observability, high-throughput data planes
Preferred skills
Large-scale inference serving experience
Technologies
Go, Kubernetes
Responsibilities
Design and develop internal scheduling system to maximize GPU utilization; Build API and tooling to manage GPU workload lifecycle; Design and improve inference control plane for faster model deployment; Collaborate with research teams to improve research velocity
Seniority
Senior, hands-on IC