Senior Manager, Cloud Services Platform
Core
Lead strategy, execution, and operations for cloud services providing container, artifact, and ML model registry capabilities for NVIDIA engineering teams.
Role type
Senior Manager, Cloud Services Platform
Builds
PaaS for GPU cloud services, container/artifact/ML model registries
Domain
Cloud infrastructure, GPU computing, developer platforms
Deliverable
production ML models | infrastructure
Required skills
Cloud platform architecture, distributed systems design, Kubernetes, containerization, artifact management, observability, team leadership, strategic planning, stakeholder management
Preferred skills
Reliability engineering (SLOs, load testing), AI infrastructure experience, regulated enterprise software delivery
Technologies
Kubernetes, Docker, Harbor, ECR, GCR, GAR, object storage, relational/NoSQL databases, event streaming
Responsibilities
Architect and operate complex PaaS for GPU cloud services; define standards for model packaging and deployment workflows; track KPIs and SLAs for registry services; mentor engineering managers and senior ICs; communicate strategy and risks to senior leadership
Seniority
Senior, hands-on IC with management responsibilities