SRE for Real-Time Streaming & Observability

Pulsar

Hong Kong

On-site

HKD 400,000 - 700,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Pulsar in Hong Kong is seeking an experienced Site Reliability Engineer focused on observability and streaming data pipelines. You will own the platform’s end-to-end monitoring, instrument services, build Grafana dashboards and alerts that clearly indicate what broke and what to check next, and participate in on-call rotations.

The role requires deploying and running production systems on Kubernetes (EKS), managing infrastructure as code with Terraform, and delivering changes through CI/CD and

Qualifications

  • University degree in CS, Software Engineering, or related field.
  • 4-7 years of relevant working experience in SRE/Observability or data platforms.
  • Experience instrumenting services, building dashboards, writing actionable alerts, and diagnosing live incidents.
  • Experience with CI/CD pipelines and GitOps practices.
  • Experience with distributed stream processing engines and OLAP databases.
  • Strong programming in at least one of Python, Scala, Java or Rust, plus SQL skills.

Responsibilities

  • Own the observability of the platform end to end: decide what to measure, instrument it, collect it, store it, and present it to people.
  • Build Grafana dashboards and alerts that name what broke and what to check next.
  • Run production systems: diagnose incidents, recover data, and fix root causes to prevent recurrence.
  • Ship changes through CI/CD and GitOps; understand the difference between merged and live.
  • Operate streaming workloads on Kubernetes (EKS) with capacity planning and config delivery.
  • Provision AWS resources behind the platform as code (Terraform).
  • Participate in on-call for the data and observability platform.
  • Build and maintain near-realtime pipelines carrying telemetry and logs to quarriable tables.
  • Reason about delivery guarantees and their impact on dashboard numbers.
  • Design data quality checks to route bad records for investigation.
  • Design and optimize OLAP tables for fast billions-row queries.
  • Write and tune SQL behind dashboards and analytics.
  • Translate requests into pipeline changes and table designs.

Skills

Observability
Grafana
Kubernetes
Terraform
SQL
Programming (Python/Scala/Java/Rust)
CI/CD
Streaming
OLAP/Columnar
On-call

Education

Bachelor's degree in Computer Science or related

Tools

Prometheus
ELK/OpenTelemetry
CloudWatch

Job description

Pulsar in Hong Kong is seeking an experienced Site Reliability Engineer focused on observability and streaming data pipelines. You will own the platform’s end-to-end monitoring, instrument services, build Grafana dashboards and alerts that clearly indicate what broke and what to check next, and participate in on-call rotations.

The role requires deploying and running production systems on Kubernetes (EKS), managing infrastructure as code with Terraform, and delivering changes through CI/CD and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, Data & Observability Platform
Site Reliability Engineer, Data & Observability Platform

Pulsar • Hong Kong

On-site
HKD 400,000 - 700,000
Site Reliability Engineer - J13203
Site Reliability Engineer - J13203

Pinpoint Asia • Hong Kong

On-site
HKD 600,000 - 900,000
Site Reliability Engineer: Observability & Data Pipelines
Site Reliability Engineer: Observability & Data Pipelines

Pinpoint Asia • Hong Kong

On-site
HKD 600,000 - 900,000
Senior SRE & Platform Innovation Lead
Senior SRE & Platform Innovation Lead

CLSA • Hong Kong

On-site
HKD 900,000 - 1,200,000
Site Reliability Engineer - HFT
Site Reliability Engineer - HFT

Selby Jennings • Hong Kong

On-site
HKD 480,000 - 720,000
Site Reliability Engineer — Observability & Data Platform
Site Reliability Engineer — Observability & Data Platform

Selby Jennings • Hong Kong

On-site
HKD 480,000 - 720,000
SRE & Observability Engineer - Reliability & Ops
SRE & Observability Engineer - Reliability & Ops

EXIO (HK) LIMITED • Hong Kong Island

On-site
HKD 450,000 - 700,000
Sr. Manager, Site Reliability & Innovation, IT
Sr. Manager, Site Reliability & Innovation, IT

CLSA • Hong Kong

On-site
HKD 900,000 - 1,200,000
Site Reliability and Observability Engineer
Site Reliability and Observability Engineer

EXIO (HK) LIMITED • Hong Kong Island

On-site
HKD 450,000 - 700,000
Mid-Level SRE
Mid-Level SRE

IO TECH SOLUTIONS LIMITED • Hong Kong

On-site
HKD 480,000 - 600,000