Senior Site Reliability Engineer (SRE)

VANGUARD SOFTWARE PTE. LTD.

Singapore

On-site

SGD 100,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Technical Leadership
Career Growth
High-Performance Team
Autonomy & Trust

Job summary

VANGUARD SOFTWARE PTE. LTD. in Singapore is seeking an experienced Senior Site Reliability Engineer (SRE) to design, build, and optimize cloud infrastructure, deployment pipelines, and observability.

You will own automation, reliability, and scalable deployments across services, guiding development teams to ship code faster and more safely. With at least 5 years in DevOps/SRE, you will lead complex incidents, shape security practices, and mentor juniors while collaborating across product teams

Qualifications

  • Bachelor's degree in Computing, Software Engineering, IT or related field.
  • Minimum 5 years of DevOps, Site Reliability Engineering (SRE), or related experience.
  • Proficient with cloud platforms (AWS, GCP, or Azure), containerization (Docker, Kubernetes), IaC (Terraform, Ansible, Helm), and CI/CD tools (Jenkins, GitHub Actions, GitLab CI/CD, ArgoCD).
  • Strong Linux administration, networking, and distributed systems knowledge.
  • Hands-on experience with monitoring/observability tools (Prometheus, Grafana, ELK/EFK, Datadog).
  • Scripting in Python, Go, Bash, etc.; ability to diagnose complex issues and design fault-tolerant systems.

Responsibilities

  • Infrastructure & Automation: Design, implement, and maintain scalable cloud infrastructure using IaC.
  • CI/CD Pipelines: Build and optimize automated pipelines for testing, deployment, and release management.
  • Monitoring & Reliability: Establish observability standards, monitoring, logging, and alerting for system health.
  • Security & Compliance: Enforce cloud security, access control, and compliance across environments.
  • Collaboration: Partner with backend, frontend, and product teams for reliable deployments.
  • Process & Mentorship: Improve DevOps processes and mentor junior engineers.

Skills

Team Mindset
Ownership
Adaptability
Communication
Problem Solving
System Design
Linux Administration
Networking
Distributed Systems

Education

Bachelor's degree in Computing/Software Engineering/IT

Tools

AWS
GCP
Azure
Docker
Kubernetes
Terraform
Ansible
Helm
Jenkins
GitHub Actions
GitLab CI/CD
ArgoCD
Prometheus
Grafana
ELK/EFK
Datadog
Python
Go
Bash

Job description

Job Summary

We are seeking a Senior Site Reliability Engineer (SRE) to join our growing engineering team. In this role, you will work independently to design, build, and optimize infrastructure and deployment pipelines that ensure the stability, scalability, and security of our systems. You will take full responsibility for automating workflows, improving observability, and enabling development teams to ship code faster and safer. This is an excellent opportunity for an experienced engineer with at least 5 years of work experience who thrives on ownership, reliability, and technical leadership.

Key Responsibilities
  • Infrastructure & Automation: Design, implement, and maintain scalable cloud infrastructure using Infrastructure as Code (IaC) tools.
  • CI/CD Pipelines: Build and optimize automated pipelines for testing, deployment, and release management.
  • Monitoring & Reliability: Establish observability standards, implement monitoring, logging, and alerting systems to ensure system health.
  • Security & Compliance: Enforce best practices for cloud security, access control, and compliance across environments.
  • Collaboration: Partner with backend, frontend, and product teams to ensure smooth deployments and reliable system operations.
  • Process & Mentorship: Improve DevOps processes, share best practices, and mentor junior engineers.
Job Requirements
  • Bachelor's Degree of Computing, Software Engineering, IT or related field.
  • Experience: Minimum 5 years of DevOps, Site Reliability Engineering (SRE), or related experience.
  • Tech Stack: Proficient with cloud platforms (AWS, GCP, or Azure), containerization (Docker, Kubernetes), IaC (Terraform, Ansible, Helm), and CI/CD tools (Jenkins, GitHub Actions, GitLab CI/CD, ArgoCD, etc.).
  • Systems Knowledge: Strong background in Linux administration, networking, and distributed systems.
  • Monitoring & Observability: Hands-on experience with tools like Prometheus, Grafana, ELK/EFK, or Datadog.
  • Scripting & Automation: Proficient in one or more languages (Python, Go, Bash, etc.).
  • Problem Solving: Skilled at diagnosing complex issues, ensuring high availability, and improving system performance.
  • System Design: Capable of designing fault‑tolerant, secure, and scalable infrastructure with disaster recovery in mind.
Soft Skills
  • Team Mindset: Collaborate effectively across teams, proactively contributing to company goals.
  • Ownership: Take responsibility for infrastructure health and ensure continuous improvements.
  • Adaptability: Open to new technologies, evolving processes, and changing business needs.
  • Communication: Clearly explain technical topics to both engineers and non‑technical stakeholders.
What We Offer
  • Technical Leadership Opportunities: Lead infrastructure design for high‑impact projects and guide DevOps best practices.
  • Continuous Growth: Access to mentorship, certifications, and a clear career progression path.
  • High‑Performance Collaboration: Work with a talented team in a modern DevOps environment (Agile/CI-CD, GitOps).
  • Flexibility and Trust: An open culture that values innovation, autonomy, and results‑driven decision‑making.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior DevOps Engineer
Senior DevOps Engineer

VANGUARD SOFTWARE PTE. LTD. • Singapore

On-site
SGD 80,000 - 120,000
Technical leadership opportunities
Access to mentorship and certifications
Modern DevOps environment
Software Engineer/ Site Reliability Engineer
Software Engineer/ Site Reliability Engineer

United States Digital Space LLC • Singapore

On-site
SGD 90,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

IDEMIA Public Security • Singapore

On-site
SGD 120,000 - 180,000
Senior Platform Engineer / Site Reliability Engineer (SRE)
Senior Platform Engineer / Site Reliability Engineer (SRE)

AMBITION GROUP SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer, Enterprise Technology Services
Site Reliability Engineer, Enterprise Technology Services

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 200,000
Head of Site Reliability Engineering (SRE) & Information Security
Head of Site Reliability Engineering (SRE) & Information Security

Kristal Advisors (Sg) Pte. Ltd. • Singapore

On-site
SGD 180,000 - 260,000
Site Reliability Engineer(Senior SRE)
Site Reliability Engineer(Senior SRE)

XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Lead Platform Site Reliability Engineer
Lead Platform Site Reliability Engineer

JPMorgan Chase & Co. • Singapore

On-site
SGD 120,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

SEVEN HILLS CONSULTING PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Purview Asia Pacific • Singapore

On-site
SGD 120,000 - 180,000