Cloud Site Reliability Engineer (SRE)

re-zoo-me

Singapore

Hybrid

SGD 120,000 - 180,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Singapore Pools (Pte) Ltd is seeking a Site Reliability Engineer to drive enterprise operational resilience by architecting scalable cloud infrastructure and managing the DevOps toolchain. You will implement GitOps, automation, and API-based services while focusing on IaC, GenAI integrations, and maximising system uptime.

Responsibilities include defining SLOs, incident response, on-call rotations, dashboards, DR planning, and FinOps reporting within a hybrid cloud environment across

Qualifications

  • Degree in Computer Science, Engineering, Information Science or related IT Discipline.
  • 4–7 years of proven hands-on software engineering, cloud, DevOps or SRE experience.

Responsibilities

  • Architect scalable hybrid cloud infrastructure with IaC and GitOps.
  • Develop automation using API-based services and GenAI integrations.
  • Define SLOs, incident response, on-call rotations, and postmortems.
  • Build centralized observability and FinOps reporting.

Skills

Python
Golang
Java
GitOps
Kubernetes
Observability

Education

Degree in Computer Science/Engineering/Information Science or related IT Discipline

Tools

Terraform
Ansible
CloudFormation
Prometheus
Grafana
Kubernetes

Job description

Date: 24 Sept 2026 Company: Singapore Pools (Pte) Ltd

Work that powers communities.

Who We Are

Singapore Pools was established by the Singapore government on 23 May 1968 to provide safe and trusted betting to counter illegal gambling. As a not-for-profit organisation, it makes contributions to the Tote Board to fund a wide range of causes in social service, community development, sports, arts, education and health sectors. Since 2004, over $5 billion have been channelled to the Tote Board. In addition, Singapore Pools also contributes about $2 billion annually to the Government in the form of taxes and duties. Its responsible gaming practices have been awarded the highest level of certification (Level 4) by the World Lottery Association's Responsible Gaming Framework since 2012. Since inception, Singapore Pools' staff have a long-standing commitment to doing good and giving back to those in need. Staff volunteers support activities held all year round, from helping disadvantaged children, youth-at-risk, underprivileged families, and elderly, to conserving the environment.

Job Purpose

The Site Reliability Engineer (SRE) drives enterprise operational resilience by architecting scalable cloud infrastructure, managing the enterprise DevOps toolchain, and ensuring centralized system observability across hybrid cloud environments. Leveraging a strong software engineering background, the SRE implements GitOps methodologies and develops custom automation applications and API-based services. By heavily utilizing Infrastructure as Code (IaC) and GenAI tools, this role eliminates manual operations, drives efficiency, optimizes cloud expenditures, and ensures maximum system uptime against stringent SLAs.

What You’ll Do
  • Software Engineering & Automation: Write code and develop automated, API-based applications (including GenAI integrations) to streamline operational reporting and eliminate recurring manual tasks.
  • Reliability & Operations: Implement SRE best practices for observability, availability, performance, and incident response. Define, measure, and govern Service Level Objectives (SLOs) and Error Budgets in collaboration with product engineering teams. Participate in on-call rotations, execute postmortem/RCAs, and identify/fix production bottlenecks.
  • Hybrid Cloud Infrastructure: Support product teams in building fault-tolerant applications by enforcing infrastructure deployment via rigorous IaC code reviews.
  • CI/CD & Toolchain: Maintain DevOps systems and enforce GitOps workflows and integrate automated security/vulnerability scanning into CI/CD pipelines for all application, container, and IaC deployment approvals.
  • Observability: Create and maintain consolidated operations dashboards by integrating telemetry from disparate monitoring tools and native AWS/Azure metrics.
  • FinOps: Generate actionable FinOps reporting to track and optimize hybrid cloud spending.
  • Resilience & Documentation: Execute disaster recovery (DR), backup, redundancy, and capacity planning strategies while maintaining high-quality runbooks and operational documentation.
Who You Are
  • Degree-qualified in Computer Science, Engineering, Information Science or related IT Discipline, with 4 to 7 years of proven, hands-on experience in software engineering, cloud architecture, DevOps, or SRE roles.
  • Hold professional certifications such as ITIL, FinOps Certified Practitioner, AWS Certified Solutions Architect (Associate or Professional), AWS Certified DevOps Engineer (Professional), AWS Certified CloudOps Engineer (Associate), Microsoft Certified Azure Administrator Associate (AZ-104), Azure Solutions Architect Expert (AZ-305), or DevOps Engineer Expert (AZ-400).
  • Cloud Architecture & Engineering: Deep hands-on experience building scalable hybrid cloud infrastructures (AWS and Azure) and containerization. Strong understanding of modern hosting, networking design patterns, and applying the Six Pillars of operational excellence across environments.
  • Software Engineering: Strong background building production-level software in Python, Golang, or Java. Experience developing and deploying API-based services and serverless applications on AWS Lambda and Azure Functions.
  • Automation & Infrastructure as Code (IaC): High proficiency in IaC and configuration management tools (e.g., Terraform, Ansible, CloudFormation) and container orchestration systems (e.g., Kubernetes). Proven capability in enforcing deployment pipelines through rigorous code review processes and working within Agile methodologies.
  • Version Control: Proficient in Git, including advanced branching strategies and GitOps paradigms.
  • Systems Knowledge: Deep expertise in architecting and managing centralized Observability platforms, utilizing GenAI, time-series databases, and diverse monitoring tools (e.g., Prometheus, Grafana, Dynatrace, Splunk, InfluxDB) to trace distributed cloud applications. Solid understanding of Database Administration and Networking architecture.
  • Professional & Interpersonal Skills: Ability to make sound, logical, data-based decisions on complex issues while considering risks. Strong communication and interpersonal skills to collaborate and build relationships with internal and external stakeholders.
  • Strong interest in technological trends and disruptions impacting Cloud Engineering and SRE.
What We Offer
  • Comprehensive total rewards package
  • Health & wellness benefits
  • Continuous learning and upskilling opportunities
  • Volunteerism and community initiatives
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Assistant Manager, OpsWatch Unit
Assistant Manager, OpsWatch Unit

Singapore Pools • Singapore

On-site
SGD 120,000 - 180,000
Total rewards
Health benefits
Learning opportunities
+1
Cloud SRE & Automation Engineer (Hybrid)
Cloud SRE & Automation Engineer (Hybrid)

re-zoo-me • Singapore

Hybrid
SGD 120,000 - 180,000
Platform Engineer (Cloud SRE Ops)
Platform Engineer (Cloud SRE Ops)

Assurity Trusted Solutions Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Annual Leave
Family Care Leave
Birthday Leave
+1
Senior Site Reliability Engineer / SRE Lead
Senior Site Reliability Engineer / SRE Lead

Reolink Technology Pte. Ltd. • Singapore

On-site
SGD 120,000 - 180,000
Insurance Coverage
Yearly Bonus & Performance Bonus
SRE Team Leader | Site Reliability Engineering
SRE Team Leader | Site Reliability Engineering

REOLINK TECHNOLOGY PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Insurance Coverage
Yearly Bonus
Performance Bonus
+1
Senior Engineer, Retail and Customer Interaction Delivery
Senior Engineer, Retail and Customer Interaction Delivery

Singapore Pools • Singapore

On-site
SGD 120,000 - 180,000
Comprehensive total rewards package
Health & wellness benefits
Continuous learning and upskilling
+1
Hybrid Cloud Operation and Delivery Engineer (SRE) - Data Infrastructure Singapore Regular
Hybrid Cloud Operation and Delivery Engineer (SRE) - Data Infrastructure Singapore Regular

ByteDance • Singapore

On-site
SGD 80,000 - 120,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Momcozy • Singapore

On-site
SGD 120,000 - 180,000
Competitive compensation
Engineer, Retail and Customer Interaction Delivery
Engineer, Retail and Customer Interaction Delivery

Singapore Pools • Singapore

On-site
SGD 60,000 - 100,000
Comprehensive total rewards package
Health & wellness benefits
Continuous learning and upskilling
+1
SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology
SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology

DBS Bank • Singapore

On-site
SGD 300,000 - 520,000