Sr Manager - Infrastructure, SRE, & AI Platforms - Services Special Projects
Core
Define long-term technical strategy and organizational structure for global, mission-critical infrastructure platforms powering consumer and enterprise workloads.
Role type
Senior Manager, Infrastructure & AI Platforms
Builds
Global compute, storage, network, observability, and AI infrastructure platforms
Domain
Cloud-native infrastructure, AI/ML, SRE
Deliverable
production ML models
Required skills
Distributed systems, Kubernetes, AI workload orchestration, Site Reliability Engineering (SRE), capacity planning, cloud architecture, data platform management, networking, security compliance, team leadership, strategic planning
Preferred skills
Large-scale enterprise infrastructure leadership, multi-engine database ecosystems, financial & capacity governance
Technologies
Kubernetes, AWS EKS, GCP GKE, Cassandra, FoundationDB, Redis, PostgreSQL, MongoDB, Kafka
Responsibilities
Define and execute long-term technical vision and capital investment strategy for global infrastructure; Architect and scale large-scale environments for AI/ML training and inference; Lead a globally distributed organization of engineers and managers; Establish SRE culture focusing on high availability and automated fault recovery; Partner with leadership to align platform capabilities with business goals.
Seniority
Senior, hands-on IC with management scope
