Senior SRE: Cloud, Automation & Incident Response

Nicoll Curtin

Singapore

Vor Ort

SGD 120.000 - 180.000

Vollzeit

Vor 2 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Nicoll Curtin is seeking a Site Reliability Engineer to support the reliability, availability and performance of production services in Central Singapore on-site for a 6–12 month renewable contract. You will monitor, troubleshoot incidents, perform RCA, and drive automation, runbook improvements and SRE practices.

The role requires 5–7 years in software engineering (JavaScript/Java/Python/.NET), 2–4 years in SRE/production support, and 3+ years AWS, Docker, Kubernetes, Terraform, plus

Qualifikationen

  • 5–7 years of software engineering experience using JavaScript, Java, Python, or .NET.
  • 2–4 years of SRE or production support experience for high-availability systems.
  • Minimum 3 years of AWS experience, including cloud services and infrastructure management.
  • Minimum 3 years of containerisation experience with Docker, Kubernetes, EKS, and Helm.
  • Hands-on experience with Terraform, CloudFormation, CI/CD workflows, GitHub Actions, and JFrog.
  • Strong Linux administration and Shell scripting capability.
  • Experience with observability and log aggregation tools such as CloudWatch, Splunk, and Datadog.
  • Sound knowledge of service monitoring, alerting, SLIs/SLOs, incident response, RCA, and post-incident actions.
  • Strong analytical, troubleshooting, written, and verbal communication skills for cross-functional collaboration.

Aufgaben

  • Monitor production services to maintain platform availability, stability, and performance.
  • Troubleshoot incidents, drive timely resolution, and communicate updates to technical and business stakeholders.
  • Perform root-cause analysis, manage post-incident follow-up, and implement reliability improvements.
  • Define and manage monitoring, alerts, SLIs, and SLOs for critical services.
  • Automate repetitive operational tasks, improve runbooks, and reduce manual support effort.
  • Diagnose complex issues across application, infrastructure, container, and cloud layers.
  • Improve developer workflows and enhance the developer experience.
  • Participate in on-call support for critical production services where required.
  • Promote best practices across software engineering, SRE, and DevOps.

Kenntnisse

JavaScript
Java
Python
.NET
SRE
AWS
Docker
Kubernetes
Terraform
CI/CD
Linux
Observability

Tools

Docker
Kubernetes
EKS
Helm
Terraform
CloudFormation
GitHub Actions
JFrog

Jobbeschreibung

Nicoll Curtin is seeking a Site Reliability Engineer to support the reliability, availability and performance of production services in Central Singapore on-site for a 6–12 month renewable contract. You will monitor, troubleshoot incidents, perform RCA, and drive automation, runbook improvements and SRE practices.

The role requires 5–7 years in software engineering (JavaScript/Java/Python/.NET), 2–4 years in SRE/production support, and 3+ years AWS, Docker, Kubernetes, Terraform, plus

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineering (Contract)
Site Reliability Engineering (Contract)

Nicoll Curtin • Singapore

Vor Ort
SGD 120.000 - 180.000
Senior Site Reliability Engineer, Cloud & Observability
Senior Site Reliability Engineer, Cloud & Observability

SCIENTEC CONSULTING PTE. LTD. • Singapore

Hybrid
SGD 120.000 - 180.000
Senior SRE – Cloud & Security Automation
Senior SRE – Cloud & Security Automation

Salt Digital Recruitment • Singapore

Vor Ort
SGD 80.000 - 120.000
Lead SRE: Cloud, Automation & High Availability
Lead SRE: Cloud, Automation & High Availability

Reolink.com • Singapore

Vor Ort
SGD 140.000 - 210.000
Insurance Coverage
Yearly Bonus
Performance Bonus
+1
Site Reliability Engineer – Cloud & Incident Response
Site Reliability Engineer – Cloud & Incident Response

XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD. • Singapore

Vor Ort
SGD 16.000 - 24.000
Health insurance
Professional development
Visa sponsorship
Senior Site Reliability Engineer — Lead Cloud & Automation
Senior Site Reliability Engineer — Lead Cloud & Automation

REOLINK TECHNOLOGY PTE. LTD. • Singapore

Vor Ort
SGD 120.000 - 180.000
5 Work Days Per Week
Office Near Tai Seng MRT
Tai Seng Exchange Tower B
+2
Senior SRE & Service Delivery Lead - Hybrid Cloud
Senior SRE & Service Delivery Lead - Hybrid Cloud

FPT Asia Pacific • Singapore

Vor Ort
SGD 90.000 - 130.000
SRE & Service Delivery Lead: Reliability & Incident Champion
SRE & Service Delivery Lead: Reliability & Incident Champion

FPT Asia Pacific Pte Ltd • Singapore

Vor Ort
SGD 120.000 - 180.000
Senior SRE: Reliability, Observability & CI/CD Lead
Senior SRE: Reliability, Observability & CI/CD Lead

NTT Data Singapore • Singapore

Vor Ort
SGD 120.000 - 180.000
SRE: Production Reliability & AI-Driven Ops
SRE: Production Reliability & AI-Driven Ops

ASTEK SINGAPORE INNOVATION TECHNOLOGY PTE. LTD. • Singapore

Vor Ort
SGD 120.000 - 180.000