VP, Site Reliability Engineer

Ambition

Singapore

On-site

SGD 250,000 - 380,000

Full time

22 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Ambition is seeking an experienced Vice President of Site Reliability Engineering to lead the reliability strategy across platforms and services in Singapore. You will establish SRE as a core engineering discipline, driving automation, observability, incident management, and reliability-by-design practices across the organization.

You will define enterprise SRE strategy, own SLIs/SLOs, and oversee incident response with a focus on scale and resilience.

Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or related field (or equivalent practical experience).
  • 12+ years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Infrastructure, DevOps, or related disciplines.
  • Proven experience building, scaling, or transforming SRE functions within complex enterprise environments.
  • Strong expertise in SRE principles and automation-first operations
  • Observability and monitoring frameworks
  • Resilience engineering and incident management
  • Capacity planning, performance optimization, and disaster recovery
  • Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budget management
  • Solid software engineering background with experience in Java and developing or supporting large-scale distributed systems.
  • Experience designing and implementing observability, automation, self-healing, or reliability platforms.
  • Strong stakeholder management skills with the ability to influence engineering teams and senior leaders without direct authority.

Responsibilities

  • Define and execute the enterprise SRE strategy, embedding reliability engineering across products, platforms, and services.
  • Own and govern SLIs, SLOs, and error budgets, ensuring reliability decisions are data-driven.
  • Drive improvements in service availability, resilience, recoverability, and performance.
  • Lead the strategy for observability, automation, self-healing capabilities, and resilience engineering platforms.
  • Oversee incident management, major incident response, post-incident reviews, and chaos engineering initiatives.
  • Drive operational excellence through automation, toil reduction, and platform standardization.
  • Partner with Engineering, Infrastructure, Security, and Product teams to embed reliability requirements early in the development lifecycle.
  • Build, mentor, and lead a high-performing team of SRE leaders and engineers.
  • Influence technical decisions across teams and stakeholders, driving reliability outcomes in a matrixed environment.

Skills

SRE principles
Observability
Incident management
Java
Distributed systems
Error budgets
Stakeholder management

Education

Bachelor's degree in Computer Science, Engineering, Information Systems, or related field

Job description

We are looking for an experienced and hands-on, Vice President of Site Reliability Engineering to lead the reliability, availability, and resilience strategy across critical platforms and services. This individual will establish SRE as a core engineering discipline, driving automation, observability, incident management, and reliability-by-design practices across the organization.

Key Responsibilities

  • Define and execute the enterprise SRE strategy, embedding reliability engineering across products, platforms, and services.
  • Own and govern SLIs, SLOs, and error budgets, ensuring reliability decisions are data-driven.
  • Drive improvements in service availability, resilience, recoverability, and performance.
  • Lead the strategy for observability, automation, self-healing capabilities, and resilience engineering platforms.
  • Oversee incident management, major incident response, post-incident reviews, and chaos engineering initiatives.
  • Drive operational excellence through automation, toil reduction, and platform standardization.
  • Partner with Engineering, Infrastructure, Security, and Product teams to embed reliability requirements early in the development lifecycle.
  • Build, mentor, and lead a high-performing team of SRE leaders and engineers.
  • Influence technical decisions across teams and stakeholders, driving reliability outcomes in a matrixed environment.

Required Qualifications

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field (or equivalent practical experience).
  • At least 12 years of experience in Site Reliability Engineering, Software Engineering, Platform Engineering, Infrastructure, DevOps, or related disciplines.
  • Proven experience building, scaling, or transforming SRE functions within complex enterprise environments.
  • Strong expertise in:
  • SRE principles and automation-first operations
  • Observability and monitoring frameworks
  • Resilience engineering and incident management
  • Capacity planning, performance optimization, and disaster recovery
  • Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budget management
  • Solid software engineering background with experience in Java and developing or supporting large-scale distributed systems.
  • Experience designing and implementing observability, automation, self-healing, or reliability platforms.
  • Strong stakeholder management skills with the ability to influence engineering teams and senior leaders without direct authority.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

VP, Site Reliability Engineer
VP, Site Reliability Engineer

AMBITION GROUP SINGAPORE PTE. LTD. • Singapore

On-site
SGD 250,000 - 350,000
Senior Vice President, Site Reliability Engineering (SRE)
Senior Vice President, Site Reliability Engineering (SRE)

Ambition Singapore • Singapore

On-site
SGD 350,000 - 520,000
Vice President, Site Reliability Engineering
Vice President, Site Reliability Engineering

Ambition • Singapore

On-site
SGD 300,000 - 550,000
Lead Platform Site Reliability Engineer
Lead Platform Site Reliability Engineer

JPMorgan Chase & Co. • Singapore

On-site
SGD 120,000 - 190,000
VP of Site Reliability & Resilience
VP of Site Reliability & Resilience

Ambition • Singapore

On-site
SGD 250,000 - 380,000
Chief SRE & Reliability Leader for Enterprise Platforms
Chief SRE & Reliability Leader for Enterprise Platforms

Ambition • Singapore

On-site
SGD 300,000 - 550,000
SL2564 - SRE & Service Delivery Lead
SL2564 - SRE & Service Delivery Lead

FPT Asia Pacific Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
SL2564 - SRE & Service Delivery Lead
SL2564 - SRE & Service Delivery Lead

FPT Asia Pacific • Singapore

On-site
SGD 90,000 - 130,000
SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology
SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology

DBS Bank • Singapore

On-site
SGD 300,000 - 520,000
VP of Reliability, Observability & SRE Automation
VP of Reliability, Observability & SRE Automation

AMBITION GROUP SINGAPORE PTE. LTD. • Singapore

On-site
SGD 250,000 - 350,000