Global VP, Reliability & Platform Observability

HP

Palo Alto (CA)

On-site

USD 320,000 - 350,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health insurance
Dental insurance
Vision insurance
Disability insurance
Employee assistance program
Flexible spending account
Life insurance
Paid parental leave
Holidays
Generous vacation

Job summary

HP is seeking a Vice President, Reliability in Palo Alto to define and lead the reliability strategy for shared platform services powering our software and experiences. You will shape the vision for highly available, observable, and scalable platforms used across HP’s products and digital experiences.

You will build a high-performing, inclusive organization of reliability, observability, performance, and resilience engineers, scale AI-enabled SRE practices, and partner with platform, product,

Qualifications

  • Bachelor’s or Master’s degree in computer science, engineering, or a related field; a Ph.D. is preferred.
  • Current or prior experience operating at the VP level in engineering within an established technology company.
  • 15+ years of experience in software, infrastructure, platform, or reliability engineering, including 8+ years leading engineering organizations through senior leadership roles.
  • Proven success building and scaling high-performing engineering teams that deliver shared capabilities used across large, complex product or platform environments.
  • Experience leading reliability or platform engineering in a global technology company operating at significant scale.
  • Experience applying AI and automation to software operations, incident response, and engineering productivity.
  • Experience supporting platforms or services that underpin products used by millions of customers and/or large-scale commercial businesses generating more than $100 million in annual revenue.
  • Deep expertise in modern cloud-native systems and platform operations, with strong experience in one or more domains: infrastructure platforms, data & AI platforms, security platforms, or developer platforms.
  • Strong track record of improving reliability, resilience, operational efficiency, and engineering effectiveness through systems thinking, automation, and disciplined execution.
  • Experience leading or scaling functions such as SRE, observability, incident management, resilience engineering, performance engineering, or large-scale service operations.
  • Demonstrated ability to influence senior executives and drive alignment across large, matrixed organizations with multiple stakeholders and competing priorities.
  • Excellent written and verbal communication skills, with the ability to communicate complex technical and operational topics clearly to executive audiences.
  • Strong collaboration and leadership presence, with a reputation for building trust, attracting top talent, and developing diverse, high-performing teams.

Responsibilities

  • Define and lead the reliability vision, strategy, operating model, and execution roadmap for shared platform services across developer experience, data and AI, security, and infrastructure.
  • Build and lead a high-performing, inclusive organization of reliability, observability, performance, and resilience leaders and engineers.
  • Establish and scale modern site reliability engineering (SRE) practices, including AI-enabled SRE, service-level objectives and indicators, error budgets, production readiness reviews, service maturity models, reliability consulting, and embedded SRE engagement models.
  • Lead the observability platform and strategy, including metrics, logs, traces, alerting, dashboards, telemetry standards, service health visibility, and developer-facing operational tooling.
  • Own incident management and resilience operations, including major incident command, escalation models, on-call standards, blameless postmortems, disaster recovery, resilience exercises, and fault-injection testing.
  • Lead performance and scalability engineering, including load testing, performance profiling, latency optimization, capacity forecasting, and scale validation for critical platform services.
  • Drive operational intelligence and automation, including operational metrics, service insights, anomaly detection, and the application of AI to improve reliability engineering and operational response.
  • Partner closely with platform, product, security, and engineering executives to embed reliability into architecture, delivery, and operations across the software lifecycle.
  • Champion a culture of accountability, engineering excellence, continuous improvement, and customer-centric decision-making.

Skills

Leadership
Reliability engineering
Cloud-native platforms
AI in operations
Incident management
Executive communication
Team building
Strategic planning

Education

Bachelor’s or Master’s in CS/Engineering
PhD preferred

Tools

Kubernetes
Cloud platforms

Job description

HP is seeking a Vice President, Reliability in Palo Alto to define and lead the reliability strategy for shared platform services powering our software and experiences. You will shape the vision for highly available, observable, and scalable platforms used across HP’s products and digital experiences.

You will build a high-performing, inclusive organization of reliability, observability, performance, and resilience engineers, scale AI-enabled SRE practices, and partner with platform, product,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

VP, Reliability & Platform Engineering
VP, Reliability & Platform Engineering

HP • Vancouver (WA)

On-site
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+5
VP, AI-Enabled Reliability & Platform
VP, AI-Enabled Reliability & Platform

Hewlett Packard Enterprise • Palo Alto (CA), Northern (KY)

Hybrid
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+7
VP, Platform Engineering & AI Excellence
VP, Platform Engineering & AI Excellence

HP • Spring (TX)

On-site
USD 280,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+3
Vice President, Reliability
Vice President, Reliability

HP • Palo Alto (CA)

On-site
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+7
Vice President, Reliability
Vice President, Reliability

Hewlett Packard Enterprise • Palo Alto (CA), Northern (KY)

Hybrid
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+7
Vice President, Reliability
Vice President, Reliability

HP • Vancouver (WA)

On-site
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+5
VP of Platform Engineering: AI, DevEx & Infra Leader
VP of Platform Engineering: AI, DevEx & Infra Leader

Hewlett Packard Enterprise • Palo Alto (CA)

On-site
USD 280,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+2
VP, Platform Engineering & Enterprise Reliability
VP, Platform Engineering & Enterprise Reliability

Future Secure AI • New York (NY)

On-site
USD 250,000 - 360,000
Flexible work environment
Competitive salary
Growth trajectory
+2
VP of Technology & Architecture — AI & Platform Strategy Leader
VP of Technology & Architecture — AI & Platform Strategy Leader

Hewlett Packard Enterprise • Palo Alto (CA), Northern (KY)

Hybrid
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+7
VP, Engineering Engagement — Internal Platform Services
VP, Engineering Engagement — Internal Platform Services

HP • Vancouver (WA)

On-site
USD 320,000 - 350,000
Health insurance
Dental insurance
Vision insurance
+7