SVP - Site Reliability Engineer

SGX Group

Singapore

On-site

SGD 320,000 - 520,000

Full time

5 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

SGX Group, a leading market infrastructure provider in Singapore, seeks an SVP of Site Reliability Engineering to shape enterprise reliability, observability, automation, and resilience across critical platforms. You will define the long-term SRE vision, collaborate with senior stakeholders, and lead senior SRE leadership teams to align reliability with business growth and regulatory expectations.

You will lead a team spanning engineering, infrastructure, security, risk, compliance, and audit,

Qualifications

  • Executive-level experience leading Site Reliability Engineering or large-scale production engineering programs.
  • Proven track record defining enterprise SRE strategy and embedding reliability standards across critical platforms.
  • Strong fluency in observability practices, SLOs/SERs, and error budgets.

Responsibilities

  • Define enterprise reliability strategy and resilience objectives.
  • Establish enterprise observability strategy and investment roadmap.
  • Define incident management framework and resilience standards.

Skills

Executive presence
Stakeholder management
Communication
Leadership
Strategic thinking
Vendor management

Education

Bachelor’s degree in Computer Science or related field

Tools

Datadog
New Relic
Grafana Cloud
AWS (multi-account)
Kubernetes/EKS
Terraform

Job description

At SGX Group, we create markets. We turn ideas into products and expand access to new opportunities. We are building the future of the exchange in-house: from architecture and platforms to the critical systems that power markets. The biggest decisions are still open to people who join us now.

Job description:

BUILD MARKETS. SHAPE ECONOMIES.

At SGX Group, we create markets. We turn ideas into products and expand access to new opportunities. We are building the future of the exchange in-house: from architecture and platforms to the critical systems that power markets. The biggest decisions are still open to people who join us now.

The opportunity

The SVP, Site Reliability Engineering is a strategic executive leadership role responsible for shaping SGX’s enterprise reliability, observability, automation, and operational resilience agenda across critical platforms and services. This leader will define the long-term SRE vision and operating model, elevate engineering and resilience standards, and ensure reliability is embedded as a strategic differentiator that supports business growth, trusted operations, regulatory confidence, and customer outcomes. Working closely with senior stakeholders across business, technology, infrastructure, security, risk, compliance, and audit, the role leads senior SRE leadership teams and drives alignment between reliability investments, critical service commitments, and enterprise transformation priorities.

SGX Ambition & Career Path:

This is a rare opportunity to shape enterprise reliability at one of Asia’s leading market infrastructures, where technology resilience, market integrity, and customer trust are inseparable. The role offers meaningful executive visibility, direct impact on business-critical outcomes, and the platform to build a world-class SRE capability in a highly regulated, innovation-driven environment.

The Site Reliability Engineering team

Site Reliability Engineering keeps the platforms behind SGX Group's business and market infrastructure available, observable, and recoverable.

The team sets the standards other engineering teams work to across service levels, error budgets, observability, incident response and automation. It works across engineering, infrastructure, security, and product, and it owns the practice as well as seen as the SME for ensuring resilience and proactive maintenance and continuous improvement of the estate.

The next phase is about establishing SRE as a discipline rather than a function, moving reliability decisions upstream into design, and reducing the manual work that currently sits behind keeping services up.

What you will do
  • Reliability Engineering: Define enterprise reliability strategy and resilience objectives.
  • Observability: Establish enterprise observability strategy and investment roadmap.
  • Incident Management: Define the enterprise incident management framework and resilience standards.
  • SLO / SLA Management: Define enterprise service performance strategy.
  • Automation: Define the function automation strategy and lead the build-out of agentic AI adoption and improvements.
  • Performance Engineering: Define the organisational scalability and performance roadmap.
  • Capacity & Resilience Planning: Own enterprise resilience and continuity planning.
  • Cloud & Platform Operations: Define cloud operations and platform reliability strategy.
  • Operational Risk Management: Help build the function operational risk strategy.
  • Engineering Leadership: Build enterprise SRE capability and workforce development strategy.
  • Agentic AI for Reliability & Operations: Help define the function strategy for AI-enabled reliability engineering, including how AI agents support resilience, incident management, observability, operational risk reduction and measurable improvements in MTTD, MTTR and service stability.
  • Partner with senior stakeholders across business, engineering, infrastructure, operations, security, risk, compliance, and audit to align reliability priorities with business strategy, regulatory obligations, and critical service commitments.
  • Provide executive leadership during major incidents and crisis scenarios, ensuring decisive cross-functional action, strong stakeholder communication, and continuous improvement.
What we are looking for
Essentials
  • Experience: Distinguished executive-level experience leading Site Reliability Engineering, platform engineering, production engineering or large-scale technology operations within complex, always-on, mission-critical environments.
  • Track record: Proven track record of defining enterprise SRE strategies, scaling leadership teams and embedding reliability standards, observability practices and engineering disciplines across critical platforms, with strong fluency in DORA, SLIs, SLOs and error budgets.
  • Technical stack: Strong experience with observability platforms such as Datadog, New Relic or Grafana Cloud, and cloud-native infrastructure including AWS multi-account, Kubernetes/EKS and Terraform.
  • Standards & environment: Experience operating in regulated, high-availability environments, with strong understanding of operational resilience, technology risk, governance, audit and compliance expectations.
  • Communication & leadership: Exceptional executive presence, stakeholder management and communication skills, with strong commercial judgement, vendor management capability and the ability to attract, inspire and retain top-tier engineering talent.
  • Education & certifications: Bachelor’s degree in Computer Science, Engineering, Information Systems, or a related discipline; advanced qualifications are advantageous.
What may set you apart
  • Exposure to financial market infrastructure, including securities and derivatives trading, clearing, settlement, market operations, or exchange-related systems.
Why this role matters

You will work on technology that underpins critical market infrastructure, where reliability is not a quality attribute of the product. It is the product. When these platforms work, participants trade and capital moves. When they do not, everyone knows within seconds.

The mandate is real. You will decide what reliability means here, how it is measured, and what the organisation is willing to trade for it. The work is demanding and the practice is still being built, which is exactly where the opportunity sits. There are not many chances in a career to establish a discipline rather than inherit one.

About SGX Group

SGX Group is one of the world's most trusted international marketplaces, known for its stability and openness. Anchored in Singapore, we enable price discovery, capital formation and risk management across asset classes, supported by resilient infrastructure and robust clearing. We convene issuers, investors and intermediaries to create and grow markets that stand the test of time. Find out more atwww.SGXGroup.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ED, Site Reliability Engineer
ED, Site Reliability Engineer

SGX Group • Singapore

On-site
SGD 250,000 - 400,000
SVP - Site Reliability Engineer
SVP - Site Reliability Engineer

Singapore Exchange Limited • Singapore

On-site
SGD 300,000 - 600,000
SVP, Enterprise SRE & Reliability Strategy
SVP, Enterprise SRE & Reliability Strategy

SGX Group • Singapore

On-site
SGD 320,000 - 520,000
ED, Site Reliability Engineer
ED, Site Reliability Engineer

Singapore Exchange Limited • Singapore

On-site
SGD 250,000 - 420,000
Executive SRE Director: Reliability, Observability & Automation
Executive SRE Director: Reliability, Observability & Automation

SGX Group • Singapore

On-site
SGD 250,000 - 400,000
SA/AVP/VP - Software Engineer
SA/AVP/VP - Software Engineer

SGX Group • Singapore

On-site
SGD 120,000 - 180,000
Executive VP, Site Reliability & Resilience
Executive VP, Site Reliability & Resilience

Singapore Exchange Limited • Singapore

On-site
SGD 300,000 - 600,000
Executive SRE & Platform Reliability Leader
Executive SRE & Platform Reliability Leader

Singapore Exchange Limited • Singapore

On-site
SGD 350,000 - 700,000
Executive Director, Site Reliability & Resilience
Executive Director, Site Reliability & Resilience

Singapore Exchange Limited • Singapore

On-site
SGD 250,000 - 420,000
Site Reliability Engineer
Site Reliability Engineer

SINGAPORE EXCHANGE LIMITED • Singapore

On-site
SGD 120,000 - 160,000