Senior Site Reliability Engineer

Quiet Capital

Jakarta Pusat

On-site

IDR 350,000,000 - 700,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Quiet Capital is seeking a Senior Site Reliability Engineer to own reliability, performance, and cost outcomes across cloud and on-prem environments. You will design Kubernetes deployment strategies, manage CI/CD pipelines, and lead infrastructure initiatives with a metrics-first approach.

You will mentor peers, drive observability improvements, and collaborate with product, security, and development teams to meet tight SLAs. Strong English communication and on-call experience are essential.

Qualifications

  • 4+ years in SRE/DevOps/Platform engineering with senior scope.
  • Deep expertise in at least one major cloud provider (AWS preferred) and Kubernetes.
  • CI/CD and GitOps (ArgoCD or equivalent) and Infrastructure as Code (Terraform).
  • Strong written and verbal English communication and documentation discipline.
  • On-call and incident handling experience during high-traffic events.

Responsibilities

  • Own availability, performance, scalability, and security of production systems end-to-end.
  • Design and evolve Kubernetes deployment strategy for production workloads.
  • Own CI/CD and GitOps pipelines and the Terraform that provisions the infrastructure behind them.
  • Diagnose and resolve database performance issues.
  • Build and maintain observability to surface problems before incidents occur.
  • Lead medium-to-large infrastructure initiatives and coordinate with stakeholders.
  • Provide technical guidance and mentorship to peers and junior engineers.

Skills

SRE/DevOps/Platform engineering
Kubernetes
Cloud computing
GitOps/ArgoCD
Database performance
Observability
Terraform
English communication

Tools

Datadog
OpenTelemetry
Terraform
ArgoCD

Job description

About The Role

The Site Reliability Engineering (SRE) team architects, builds, and maintains the rock-solid infrastructure that applications rely on. At the Senior Level, you own reliability, performance, and cost outcomes for the systems under your area end-to-end, not just executing well-defined tasks, but deciding between tradeoffs, scoping ambiguous problems, and driving process and system improvements that span teams. You'll work closely with development, security, and product teams, and mentor other engineers as a technical point of reference for the team.

What You Will Do
  • Own the availability, performance, scalability, and security of production systems end-to-end, across cloud (AWS/GCP) and on-premises environments.
  • Design and evolve Kubernetes deployment strategy for production workloads.
  • Own CI/CD and GitOps pipelines in production (ArgoCD or equivalent) and the Terraform that provisions the infrastructure behind them.
  • Diagnose and resolve database performance issues.
  • Build and maintain observability that surfaces problems before they become incidents.
  • Seek out and implement process and system improvements affecting performance and security, coordinating with multiple stakeholders.
  • Scope and lead medium-to-large infrastructure initiatives: gather requirements, prioritize by business impact, and communicate impact to stakeholders.
  • Negotiate technical tradeoffs with stakeholders to meet business SLAs.
  • Provide technical guidance and mentorship to peers and junior engineers; promote best practices and standards across the team.
  • Maintain documentation and process discipline for the systems and incidents you own.
  • Lead structured incident investigation, isolating server, database, and application layers with a metrics-first approach, including on-call during high-traffic events.
What Are We Looking For
  • At least 4 years of experience in either SRE, DevOps, MLOps, or platform engineering, including senior-level scope at a high-traffic company.
  • Deep expertise in one major cloud provider (preferably AWS), with a proven ability to ramp up on the other quickly.Production experience & expertise with Kubernetes & Linux fundamentals
  • CI/CD & GitOps (ArgoCD or other equivalent stacks)
  • Database performance analysis & monitoring (MySQL, Postgres)
  • Observability tooling & standards (Datadog, OpenTelemetry)
  • Infrastructure as Code (Terraform)
  • Strong working English, verbal & written communication. Strong documentation and process discipline.
  • On-call & incident handling experience during high-traffic events.
  • Comfortable negotiating with stakeholders to propose technical compromises that meet business SLAs.
  • Demonstrated growth mindset and proven ability to own ambiguous scope.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

StraitsX • Jakarta Pusat

On-site
IDR 600,000,000 - 840,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

StraitsX • Kebayoran Baru

On-site
IDR 600,000,000 - 1,000,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

StraitsX • Indonesia

On-site
IDR 300,000,000 - 540,000,000
Senior Site Reliability Engineer New Jakarta, Jakarta, Indonesia
Senior Site Reliability Engineer New Jakarta, Jakarta, Indonesia

StraitsX Group • Jakarta Pusat

On-site
IDR 550,000,000 - 750,000,000
Site Reliability Engineer
Site Reliability Engineer

PARTECH PARTNERS • Daerah Khusus Ibukota Jakarta

On-site
IDR 272,380,000 - 453,968,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

AccelByte • Kota Yogyakarta

On-site
IDR 2,144,389,000 - 3,216,583,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BookCabin • Jakarta Pusat

On-site
IDR 250,000,000 - 420,000,000
Senior SRE: Scale, Reliability & Automation Leader
Senior SRE: Scale, Reliability & Automation Leader

StraitsX • Jakarta Pusat

On-site
IDR 600,000,000 - 840,000,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

AccelByte • Sleman

On-site
Senior SRE: End-to-End Reliability & Platform Lead
Senior SRE: End-to-End Reliability & Platform Lead

StraitsX Group • Jakarta Pusat

On-site
IDR 550,000,000 - 750,000,000