Site Reliability Engineer — Scale Data Infra on AWS & ClickHouse

PostHog

United States

On-site

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

PostHog is a fully remote product analytics company seeking a Senior Site Reliability Engineer to own production infrastructure at scale on AWS. You’ll manage large fleets of EC2-based VM workloads, improve deploy tooling, and collaborate with ClickHouse engineers to turn database needs into infra solutions.

The role emphasizes end-to-end ownership, on-call readiness, and building self-healing automation to reduce incidents.

Qualifications

  • Prior experience with ClickHouse or other OLAP databases.
  • Strong experience operating production infrastructure on AWS.
  • Hands-on experience with VM-based systems (EC2).
  • Experience automating infrastructure using Terraform, Ansible, or similar.
  • Solid understanding of Linux systems (disk, memory, networking, failure modes).
  • Experience supporting stateful systems (databases, queues, storage systems, etc.).
  • Ability to debug and reason about performance and reliability issues in production.
  • You’re comfortable owning systems end-to-end, including on-call responsibilities.

Responsibilities

  • Managing large fleets of EC2-based VMs, disks, and networking for data-intensive workloads.
  • Improving operational tooling around deploys, schema changes, backups, restores, and incident response.
  • Working closely with ClickHouse engineers to turn database-level needs into infra-level solutions.
  • Reducing operational load by identifying repeat pain points and eliminating them through code and self-healing automation.
  • Participating in on-call and incident response, with a strong focus on making incidents rarer over time.
  • You’ll have room to design and automate, not just respond to alerts.

Skills

ClickHouse
OLAP databases
AWS
Linux
Stateful systems
On-call ownership

Tools

Terraform
Ansible
EC2

Job description

PostHog is a fully remote product analytics company seeking a Senior Site Reliability Engineer to own production infrastructure at scale on AWS. You’ll manage large fleets of EC2-based VM workloads, improve deploy tooling, and collaborate with ClickHouse engineers to turn database needs into infra solutions.

The role emphasizes end-to-end ownership, on-call readiness, and building self-healing automation to reduce incidents.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ClickHouse Infrastructure Engineer: Scale & Automation
ClickHouse Infrastructure Engineer: Scale & Automation

PostHog • United States

On-site
USD 150,000 - 210,000
ClickHouse Infrastructure Engineer
ClickHouse Infrastructure Engineer

PostHog • San Francisco (CA)

On-site
USD 150,000 - 210,000
Senior SRE — Remote, Cloud Reliability & Uptime
Senior SRE — Remote, Cloud Reliability & Uptime

ClickHouse • United States

Remote
USD 150,000 - 210,000
Flexible work environment
Healthcare
Stock options
+3
Remote SRE – Scale, Automation & Ownership (US CT/ET)
Remote SRE – Scale, Automation & Ownership (US CT/ET)

PostHog • United States

On-site
USD 140,000 - 210,000
Senior SRE — Postgres Infra for Cloud Data Platform
Senior SRE — Postgres Infra for Cloud Data Platform

ClickHouse, Inc. • Northern (KY)

Hybrid
USD 150,000 - 210,000
Flexible work environment
Healthcare
Equity in the company
+3
ClickHouse Operations Engineer
ClickHouse Operations Engineer

PostHog • United States

On-site
USD 150,000 - 210,000
Senior ClickHouse Data Architect, AWS Lakehouse
Senior ClickHouse Data Architect, AWS Lakehouse

Careers at Lineate • Georgia

On-site
USD 140,000 - 180,000
Freedom to Develop
Career path and performance management
Social benefits package
+6
Senior DB Reliability Engineer - Remote & Flexible Hours
Senior DB Reliability Engineer - Remote & Flexible Hours

Alex Staff • Town of Poland (NY)

On-site
USD 130,000 - 190,000
Fully remote work
Flexible working hours
Vacation 24 days per year
+5
Remote SRE/DevOps Engineer - Kubernetes, AWS, ClickHouse
Remote SRE/DevOps Engineer - Kubernetes, AWS, ClickHouse

Ekloud Data Labs • United States

Remote
USD 96,000 - 165,000
Remote work
Contract role
Remote Cloud-Native Data Infrastructure Engineer
Remote Cloud-Native Data Infrastructure Engineer

Empleora • Northern (KY)

Hybrid
USD 133,000 - 197,000