Vice President, Reliability
Core
Define and lead the reliability strategy, operating model, and execution for shared platform services powering HP's products and digital experiences to ensure high availability, resilience, and performance at global scale.
Role type
Senior leadership VP of Reliability (SRE/Platform Operations)
Builds
Shared platform services for developer experience, data & AI, security, and infrastructure
Domain
Technology / Cloud-native Platform Engineering / Reliability Engineering
Deliverable
production ML models | infrastructure | dashboards & analysis
Required skills
Reliability strategy definition, SRE practices (error budgets, service maturity), observability platform leadership, incident management & resilience operations, performance & scalability engineering, operational intelligence & automation, systems thinking, executive influence, team building & scaling
Preferred skills
AI/ML application to operations, PhD in CS/Engineering
Technologies
Cloud-native systems, observability tools (metrics, logs, traces), AI/ML frameworks for operations
Responsibilities
Define reliability vision and roadmap for shared platforms; Build and lead high-performing reliability organizations; Establish modern SRE practices and observability standards; Own incident management, disaster recovery, and resilience exercises; Lead performance engineering and capacity planning; Drive operational automation and AI integration; Partner with executives to embed reliability in the software lifecycle
Seniority
VP level, strategic leadership with hands-on technical direction