Cloud Observability Engineer: Scale & On-Call Reliability

ClickHouse

San Francisco (CA)

Remote

USD 180,000 - 260,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Flexible work environment
Healthcare
Equity in the company
Flexible time off
Home office stipend

Job summary

ClickHouse is seeking an experienced software engineer to join the Observability organization. You will design, build, and operate distributed telemetry systems at scale and help own the reliability and performance of our telemetry pipelines and storage layers.

You will work across Observability Platform and Internal Observability, collaborate with product and infra teams, and participate in on-call rotations to fix incidents and improve tooling. Remote-friendly with global opportunities.

Qualifications

  • 5+ years of experience building and operating production systems at scale.
  • Strong proficiency in Go.
  • Experience building and operating services on Kubernetes.
  • Production experience with a major cloud provider (AWS/GCP/Azure).
  • Hands-on experience with telemetry systems such as OpenTelemetry, Prometheus, Grafana.

Responsibilities

  • Design, build, and operate distributed systems that ingest, process, and store telemetry at very high scale.
  • Own the reliability, performance, capacity, and cost-efficiency of telemetry pipelines and storage systems.
  • Participate in on-call rotations, resolve production incidents, and drive root-cause fixes to completion.
  • Build software and automation to eliminate repetitive operational work and simplify platform operations.
  • Identify bottlenecks and help shape the roadmap for future scale.

Skills

Go
Kubernetes
Distributed systems
Cloud (AWS/GCP/Azure)
Telemetry

Tools

Terraform
Helm
Argo CD
OpenTelemetry
Prometheus
Grafana
GitOps

Job description

ClickHouse is seeking an experienced software engineer to join the Observability organization. You will design, build, and operate distributed telemetry systems at scale and help own the reliability and performance of our telemetry pipelines and storage layers.

You will work across Observability Platform and Internal Observability, collaborate with product and infra teams, and participate in on-call rotations to fix incidents and improve tooling. Remote-friendly with global opportunities.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Cloud Observability Engineer - Scale Telemetry
Remote Cloud Observability Engineer - Scale Telemetry

ClickHouse, Inc. • Northern (KY)

Hybrid
USD 140,000 - 190,000
Flexible work environment
Healthcare
Equity in the company
+3
Senior Backend Engineer Observability Platform (Remote)
Senior Backend Engineer Observability Platform (Remote)

ClickHouse, Inc. • Northern (KY)

Hybrid
USD 120,000 - 180,000
Flexible work environment
Healthcare
Equity in the company
+3
Cloud Software Engineer - Observability Platform
Cloud Software Engineer - Observability Platform

ClickHouse • San Francisco (CA)

Remote
USD 180,000 - 260,000
Flexible work environment
Healthcare
Equity in the company
+2
Senior Full-Stack Engineer — Remote Observability Platform
Senior Full-Stack Engineer — Remote Observability Platform

ClickHouse, Inc. • Northern (KY)

Hybrid
USD 120,000 - 180,000
Flexible work environment
Healthcare
Equity in the company
+3
Senior Backend Engineer — Observability Platform (Remote)
Senior Backend Engineer — Observability Platform (Remote)

ClickHouse • United States

Remote
USD 120,000 - 160,000
Flexible work environment
Employer contributions towards healthcare
Equity in the company
+2
Cloud Software Engineer - Observability Platform
Cloud Software Engineer - Observability Platform

ClickHouse, Inc. • Northern (KY)

Hybrid
USD 140,000 - 190,000
Flexible work environment
Healthcare
Equity in the company
+3
Senior Observability Engineer — Scale & Reliability Leader
Senior Observability Engineer — Scale & Reliability Leader

Temporal Technologies • United States

On-site
USD 141,000 - 231,000
Senior Backend Engineer for Observability Platform (Remote)
Senior Backend Engineer for Observability Platform (Remote)

ClickHouse • San Francisco (CA)

Remote
USD 150,000 - 190,000
Flexible work environment
Healthcare
Equity in the company
+3
Observability Engineer: Scale Telemetry & AI Reliability
Observability Engineer: Scale Telemetry & AI Reliability

Baseten • New York (NY)

On-site
USD 165,000 - 330,000
Competitive compensation including 1x/
Full medical/dental/vision insurance
Flexible PTO with Winter Break
+3
Senior Site Reliability Engineer - Cloud Reliability (Remote)
Senior Site Reliability Engineer - Cloud Reliability (Remote)

ClickHouse • United States

Remote
USD 140,000 - 210,000
Flexible work environment
Healthcare
Equity in the company
+3