Eine zielgenaue Bewerbung für diesen Job — ein maßgeschneiderter Lebenslauf und ein Anschreiben, die genau zur Stellenanzeige passen.
Fact Finder in Berlin is seeking an experienced SRE to help build and shape a modern on-prem private cloud platform. You will own reliability topics, craft SLOs/SLIs, and work closely with developers to ensure fast, reliable product delivery.
You’ll lead incident response, drive automation, and evolve observability across two stacks, with a plan for hybrid cloud expansion and strong emphasis on collaboration and learning.
Location & work model : Berlin, hybrid
Tech stack: Kubernetes on our own servers, Harvester ( KubeVirt ), Argo CD/Flux, Prometheus/Grafana, Longhorn/Ceph
Team: A growing SRE team - you report to our CTPO for now and to the Team Lead SRE we're hiring next; two system administrators in Pforzheim run the physical hardware
Process: Intro call - take-home task (~2h) - 90-min tech interview with our developers - leadership conversation - meet the team
Languages: Fluent English required; German is a plus, not a must
Most SRE jobs today mean clicking around a managed cloud console. This one doesn't. We run our own hardware in Frankfurt and are building a modern private cloud platform on Kubernetes and Harvester - on-prem by default, with elastic burst into the public cloud and the option to go cloud-only later. You won't inherit a finished SRE practice: you'll help define it, side by side with our Berlin development teams - and you won't do it alone, a Team Lead SRE hire is coming next. SRE here is an enabling discipline: you build what our developers need to ship reliably, while two system administrators in Pforzheim run the physical hardware. And the impact is direct - our product discovery technology powers more than 2,000 European online shops (Intersport, SPAR, Douglas and more), handling billions of shopper queries a year. When product discovery is slow or down, our customers lose revenue in real time.
You get to know both products, join the on-call rotation with a buddy, and own your first reliability topic - SLOs for one product, alerting that actually helps at 3 a.m., or automating away a piece of toil. By day 90 you've shipped visible improvements and know where you want to take the platform next.
You don't tick every box - or your title was never "SRE"? Apply anyway. If you've owned production systems, handled incidents and worked deeply with Kubernetes, we want to hear from you - production experience and engineering mindset matter more to us than titles or buzzwords.
Impact from day one: Your work directly influences the revenue of leading eCommerce brands across Europe.
Modern tech stack: Kubernetes, Harvester, GitOps, auto-scaling, and an exciting path toward the cloud - with room to build things right.
AI-first mindset: We use AI as a real part of our daily work, not as a buzzword.
Ownership & growth: Clear responsibility, short decision paths, and the opportunity to actively shape your role.
Flexible work: Hybri