Erhalte mehr Antworten von Arbeitgebern
Versende in nur wenigen Minuten einen passgenauen Lebenslauf.
PostHog is seeking SREs to take deep ownership of production systems, especially a petabyte-scale ClickHouse deployment on AWS. The role centers on turning a fast-growing, stateful platform into a reliable, automated system, with provisioning, scaling, recovery, and self-healing automation at its core.
You'll work across databases and infra, reducing operational stress and building patterns that scale without proportionally increasing human effort.
We’re looking for people (EU/UK based) that like deep ownership of production systems, people that are not afraid of working with stateful infrastructure and love working in AWS, VMs, automation, and making messy systems reliable.
We run one of the largest self-managed ClickHouse installations on AWS, at petabyte scale, and we’re actively preparing it for the next 10–50× of growth. This role sits at the centre of that effort. You won’t be in a typical "keep the lights on" SRE role. The work is about turning a fast-growing, stateful system into a predictable, well-automated platform (provisioning, scaling, rebalancing, recovery). That means reducing operational stress, designing safe automation for data‑heavy workloads, and building the tooling and patterns that let the system scale without scaling human effort. You’ll work on the kind of problems that only show up at large scale (petabytes of data, thousands of cores, constant ingestion).
You should join this team if you like deep ownership of production systems, and are not afraid of working with stateful infrastructure.
You don’t need to be a ClickHouse expert on day one. We’ll teach you the database internals, but you do need to enjoy owning complex infrastructure.
We are committed to ensuring a fair and accessible interview process. If you need any accommodations or adjustments, please let us know.