Transforma esta oferta en una entrevista — un currículum y una carta de presentación creados pensando en lo que quiere el empleador.
ClickHouse, in partnership with Langfuse Cloud, seeks an experienced SRE to run production infrastructure at scale. You will own deployments on AWS ECS Fargate and ClickHouse Cloud, manage autoscaling and cost control, and advance our self-hosted offerings with Helm and Docker Compose documentation.
You’ll build robust monitoring, dashboards, and alerts in Datadog, drive automation across CI/CD, and contribute to security and reliability in a small, accountable team.
Your work will keep Langfuse running — everywhere.
Langfuse processes over a billion trace events per month. When a Fortune 50 company relies on Langfuse in production, they're relying on the infrastructure you operate. You'll own uptime, performance, and cost efficiency across our entire cloud footprint — and you'll make sure every self-hosted deployment runs just as smoothly.
You’ll operate Langfuse Cloud on AWS ECS Fargate and ClickHouse Cloud, with Datadog as the observability backbone. You’ll also own our public self-hosted infrastructure — including our Helm chart, Docker Compose setup, and everything in between — so that teams from startups to enterprises can run Langfuse on their own terms.
This isn’t a "maintain what exists" role. We’re scaling fast, and you’ll be the person who makes sure the infrastructure grows ahead of demand — not behind it.
Langfuse is now part of ClickHouse, which means the team behind the database at the core of our stack is one channel away. Few infrastructure roles give you that kind of direct access to the people who build your most critical dependency.
You’ll run our production environments on AWS ECS Fargate and ClickHouse Cloud. You’ll manage deployments, autoscaling, capacity planning, and cost optimization — making sure we stay fast and affordable as traffic scales.
You’ll own our Datadog setup end to end — dashboards, alerts, and SLOs. When something degrades, you’ll ensure we know before our customers do. You’ll build the monitoring culture that lets the whole team ship with confidence.
Thousands of teams run Langfuse on their own infrastructure. You’ll own and evolve our Helm chart, Docker Compose configuration, and deployment documentation. You’ll turn "works on my machine" into "works on every machine" — from a single-node setup to a multi-region enterprise deployment.
CI/CD pipelines, infrastructure‑as‑code, automated scaling, zero‑downtime deployments. You’ll replace manual processes with automation that makes the team faster and the platform more reliable.
We’re growing fast and new product directions — like complex long‑running agent observability and real‑time evaluation — push the infrastructure in new ways. You’ll be thinking ahead about what breaks at 10x scale and building the foundation before we get there. 10x is always just one quarter away here at Langfuse.
As more enterprises adopt Langfuse, you’ll help ensure our cloud and self‑hosted deployments meet the security and compliance bar that large organizations require.
Culture - We All Shape It
As part of a rapidly scaling start‑up, you will be instrumental in shaping our culture.
Are you interested in finding out more about our culture? Learn more about our values here. Check out our blog posts or follow us on LinkedIn to find out more about what’s happening at ClickHouse.
ClickHouse provides equal employment opportunities to all employees and applicants and prohibits discrimination and harassment of any type based on factors such as race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.
Please see here for our Privacy Statement.
Compensation Range: €90K - €160K