Senior Systems Software Engineer - NV Cloud Functions
Core
Build and evolve open-source, cloud-native software infrastructure to deploy, manage, and serve GPU-accelerated AI workloads across distributed regions and clusters.
Role type
Senior Systems Software Engineer (Distributed GPU Infrastructure)
Builds
Scalable services for inference, streaming, and batch workloads on distributed GPU fleets
Domain
Cloud-native AI infrastructure, distributed systems, GPU/DPU acceleration
Deliverable
production ML models | infrastructure
Required skills
Systems programming (Go, C, Rust), Kubernetes, container orchestration, CI/CD automation, Linux/Unix internals, distributed systems architecture, performance engineering, scripting (Bash/Python)
Preferred skills
Publish-subscribe architectures, message queues, HTTP/2 and gRPC optimization, Kubernetes Custom Resources and Operators
Technologies
Java, Go, Rust, Kubernetes, GitLab, ArgoCD, gRPC, HTTP/2
Responsibilities
Design and ship scalable services for distributed GPU workloads; optimize performance and reliability of workload routing systems; develop control-plane and edge components; automate build, testing, and deployment processes; contribute to public open-source repositories and triage community issues; investigate complex distributed infrastructure challenges.
Seniority
Senior, hands-on IC