Senior Site Reliability Engineer - Scale & Resilience Leader

Kidentify

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Kidentify seeks a Senior Site Reliability Engineer to own and improve the reliability of our platform. You will manage infrastructure on AWS and Kubernetes, build strong observability, and lead incident response for a globally scaled service.

You will collaborate with engineering to bake reliability into design, automate toil away, and optimize deployment safety with robust CI/CD and environment controls. This is a hands-on role requiring coding and systems thinking.

Qualifications

  • 5+ years of production ownership in infrastructure, platforms, SRE, or software with ownership responsibilities.
  • Strong experience running production systems on AWS.
  • Hands-on experience with Kubernetes and container workloads.
  • Experience with infrastructure as code, preferably Terraform.
  • Experience designing observability stacks with Prometheus, Grafana, OpenTelemetry, or equivalents.
  • Strong understanding of distributed systems, failure modes, and production debugging.
  • Experience building or improving CI/CD systems and release workflows.
  • Ability to write code and automation in Go, Python, or TypeScript.
  • Good judgment during incidents with a practical mindset on tradeoffs and recovery.

Responsibilities

  • Own the reliability, availability, and performance of the systems behind k-ID's platform and public APIs.
  • Design and improve scalable infrastructure on AWS and Kubernetes to support growth and global workloads.
  • Build and maintain observability across logs, metrics, tracing, alerting, and service health.
  • Improve deployment safety through CI/CD workflows, release controls, and environment consistency.
  • Lead incident response and production readiness practices, including runbooks, on-call hygiene, and postmortems.
  • Reduce operational toil by automating repetitive tasks and improving internal tooling.
  • Partner with engineering teams to embed reliability and operability from the start of service design.
  • Strengthen platform security and hygiene across access controls, secrets handling, and hardening.
  • Continuously improve system performance and cost awareness without sacrificing reliability.

Skills

AWS
Kubernetes
CI/CD
Go
Python
TypeScript
Observability
Distributed systems
Incident response

Tools

Terraform
Prometheus
Grafana
OpenTelemetry

Job description

Kidentify seeks a Senior Site Reliability Engineer to own and improve the reliability of our platform. You will manage infrastructure on AWS and Kubernetes, build strong observability, and lead incident response for a globally scaled service.

You will collaborate with engineering to bake reliability into design, automate toil away, and optimize deployment safety with robust CI/CD and environment controls. This is a hands-on role requiring coding and systems thinking.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kidentify • Singapore

On-site
SGD 120,000 - 180,000
Lead Platform Site Reliability Engineer
Lead Platform Site Reliability Engineer

JPMorgan Chase & Co. • Singapore

On-site
SGD 120,000 - 190,000
Senior Platform Engineer / Site Reliability Engineer (SRE)
Senior Platform Engineer / Site Reliability Engineer (SRE)

Ambition • Singapore

On-site
SGD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

TP-LINK CORPORATION PTE. LTD. • Singapore

On-site
SGD 90,000 - 150,000
Senior Site Reliability Engineer, Cloud & Observability
Senior Site Reliability Engineer, Cloud & Observability

SCIENTEC CONSULTING PTE. LTD. • Singapore

Hybrid
SGD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

SEVEN HILLS CONSULTING PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
Senior Platform Engineer / Site Reliability Engineer (SRE)
Senior Platform Engineer / Site Reliability Engineer (SRE)

AMBITION GROUP SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
Cloud-Scale SRE: Reliability, Observability & Incident Mastery
Cloud-Scale SRE: Reliability, Observability & Incident Mastery

SEVEN HILLS CONSULTING PTE. LTD. • Singapore

On-site
SGD 90,000 - 130,000
AVP Site Reliability Platform Lead
AVP Site Reliability Platform Lead

Singapore Exchange Limited • Singapore

On-site
SGD 140,000 - 210,000
Site Reliability Engineer
Site Reliability Engineer

RemotePeople • Singapore

On-site
SGD 120,000 - 190,000