Senior/Staff Platform Engineer (m/f/x)

DUDE CHEM

Berlin

Vor Ort

EUR 120.000 - 180.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Cortea is building a platform that sits between software engineering, DevOps and SRE, with customers who are other engineers. The role focuses on taking systems from design to production, owning the architecture, and ensuring reliability across a growing Kubernetes and cloud footprint.

You will design, implement, and defend scalable foundations, provide the tooling and documentation that engineers rely on, and drive security and compliance through trusted processes.

Qualifikationen

  • Designed, built and operated distributed systems end to end.
  • Strong backend software engineering and DevOps/SRE experience.
  • Experience running production Kubernetes and large-scale observability.
  • Scaled PostgreSQL under real load and managed performance under stress.

Aufgaben

  • Own the architecture and core stack, codify golden paths and APIs.
  • Lead infrastructure, observability, and reliability initiatives with SRE focus.
  • Define SLIs/SLOs, and drive CI/CD and cloud footprint improvements.
  • Build AI development tooling and documentation to keep engineers on the paved road.

Kenntnisse

Distributed systems
Backend engineering
DevOps/SRE
Kubernetes
PostgreSQL
Observability
Design docs

Tools

Kubernetes
CI/CD pipelines
Cloud platforms
Observability tooling

Jobbeschreibung

We ship fast, and we intend to keep shipping fast. The consequence is that our system is outgrowing the amount of deliberate design that went into it. We run AI agents over customer documents in a domain where being wrong is expensive, and a lot of the load-bearing decisions about how those fit together are still implicit.

You will be making those decisions explicit and then making them hold. That means codifying golden paths, building the shared primitives and APIs that product engineers work on top of, owning the architecture of the system as a whole, and establishing what our reliability actually needs to be before an incident establishes it for us.

Platform at Cortea sits between software engineering, DevOps and SRE. Its customers are other engineers. This is not a DevOps role under a different name: you will spend more time in application code than in YAML, and the reason we want infrastructure experience is that we don't believe you can design a system well without being able to run it.

The bar we are hiring against is someone who has taken a system from nothing to production and owned it end to end. You picked the technology, argued the tradeoffs in writing, provisioned the infrastructure, and were responsible for it when it broke.

On AI

We use AI heavily and we want you to. What we are not looking for is someone who outsources their judgment to it.

The mental model of this system has to live in your head, not in a context window. Use agents to move quickly on implementation. Do the design, the reasoning and the writing yourself, and be able to defend every decision you ship.

Four areas, roughly in the order you will spend time on them. You won't work on all of them at once, but you should be open to any of them.

  • Product platform. Make the easiest way for product engineers to do something (the paved road, or golden path) also the most secure, reliable and scalable way by default. Build the shared primitives, libraries and APIs that hide complexity and carry our quality and observability standards with them. Own the architecture and the core stack: what gets standardized, what gets reused, and where the system should be in eighteen months.

  • Infrastructure, observability and reliability. Infrastructure as code by default, from cloud resources through to dashboards and alerts. Provision, run and tune our Kubernetes cluster and cloud footprint, and keep CI fast as the deploy rate grows. Define SLIs for the workloads that matter, build SLOs on top of them, and make the alerting high-signal enough that people trust it. Drive down AI and infrastructure spend.

  • Security and compliance. Keep the internal foundations secure by default: IAM, dependency management, secrets. Own the authentication and authorization stack, including ReBAC models covering both humans and agents. Implement SOC 2 and ISO 27001 controls without taxing every future change.

  • AI dev tooling. Keep product engineers and their agents on the paved road by making sure the documentation and agent guidelines they need are in place. Shorten the path from design to implementation with standardized automations and development environments that stay close to production.

What this looks like in practice
  • A system of record for managing and distributing constantly evolving agent configurations, under strict auditability and tenant isolation requirements

  • The harnesses powering our AI agents, abstracting over use cases that change faster than the code

  • Centralized progress tracking that holds up across a growing number of parallel agent executions

  • Usage-based tenant billing with quotas and rate limiting across a growing set of products

  • SLIs for the background job system behind our agents (execution latency, AI spend, memory, CPU), then fixing the bottlenecks they expose at both the application and infrastructure level

  • A document pipeline handling dozens of file formats, very large spreadsheets and hundreds of parallel uploads, under a strong reliability requirement

You will fit into this role if you…

  • Have designed, built and operated distributed systems end to end, and enjoy understanding how every part interacts with the rest

  • Are strong at backend software engineering and at DevOps/SRE, and don't think of those as separate jobs

  • Write design docs, RFCs, ADRs, postmortems

  • Reach for simple, boring solutions first, can tell essential complexity from accidental, and know which corners are safe to cut and which ones compound

  • Can turn a hard problem on its head and find the alternative nobody proposed

Concretely, you have scaled Postgres or another relational transactional database under real load, run production Kubernetes, and built observability rather than inherited it: SLIs, SLOs, distributed tracing.

No one checks every box. If you have designed a system from scratch, run it in production, and can explain in writing why it is shaped the way it is, let's talk.

Nice-to-haves that are a plus:
  • Durable workflow orchestrators like Temporal, and background job and queue processing generally

  • Infrastructure for LLM-based products or agentic systems

  • Audit, finance, compliance, or another high-accuracy domain

  • Azure and/or GCP

This is probably not the right role if…
  • You would describe application code as someone else's responsibility. A large share of platform work here happens inside the product codebase.

  • Your answer to reliability is more process. We want guardrails, not gates.

  • You need a well-defined system to work in. Ours is not one yet, and shaping it is the job.

What you will get

  • You own the brand and shape how the product feels. Work face-to-face with experienced founders and learn directly from customer insights.

  • Real influence on strategy. In a small team of excellent engineers and operators with high autonomy, you will shape product direction. Best idea wins, regardless of seniority.

  • Fast, ambitious, and fun team. Decisions get made in hours, not weeks. Rapid experimentation is the default.

  • Meaningful equity and competitive salary. You are one of the first to build this, and you share in the upside.

  • A mission that matters. Building intelligent systems for a $200bn industry. From Berlin, with AI at the core.

Interview process

  • First Call — Intro to Cortea with Liza

  • Second Call — Technical interview with Dan

  • Third Call — Deep dive into our culture with our Co-Founder Philipp

  • On-site Day (Berlin) — Meet the team and work on a real problem together

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior/Staff Platform Engineer (m/f/x)
Senior/Staff Platform Engineer (m/f/x)

United States Digital Space LLC • Berlin

Vor Ort
EUR 110.000 - 160.000
Equity
Senior/Staff Fullstack Engineer (m/f/x)
Senior/Staff Fullstack Engineer (m/f/x)

Cortea • Berlin

Hybrid
EUR 90.000 - 130.000
High impact & growth
Competitive salary + equity
Learning budget for courses and confs
+1
Senior/Staff Platform Engineer (m/f/x)
Senior/Staff Platform Engineer (m/f/x)

Meyandy LLC • Berlin

Hybrid
EUR 110.000 - 170.000
AI Product Manager (m/f/x)
AI Product Manager (m/f/x)

Cortea • Berlin

Vor Ort
EUR 90.000 - 130.000
High impact & growth
Mission-driven culture
Competitive salary
+2
Senior/Staff AI Engineer, Quality & Evals (m/f/x)
Senior/Staff AI Engineer, Quality & Evals (m/f/x)

Cortea AI • Berlin

Vor Ort
EUR 110.000 - 170.000
Significant equity
Generous tools budget
Flexible vacation
+3
Chief of Staff (m/f/x)
Chief of Staff (m/f/x)

Cortea AI • Berlin

Vor Ort
EUR 90.000 - 160.000
Learning budget
Team lunches
Off-sites
+2
Chief of Staff (m/f/x)
Chief of Staff (m/f/x)

Cortea • Berlin

Vor Ort
EUR 90.000 - 150.000
Learning budget
Team lunches
Off-sites
+2
AI Engineer, Quality & Evals (m/f/x)
AI Engineer, Quality & Evals (m/f/x)

Jackalope Digital LLC • Berlin

Vor Ort
EUR 90.000 - 130.000
High impact & growth
Mission-driven culture
Attractive compensation
+3
AI Engineer, Quality & Evals (m/f/x)
AI Engineer, Quality & Evals (m/f/x)

Cortea AI • Berlin

Vor Ort
EUR 90.000 - 140.000
Equity
Flexible vacation
Team lunches
+2
Principal Product & Brand Designer (m/f/x)
Principal Product & Brand Designer (m/f/x)

Cortea-Ai • Berlin

Vor Ort
EUR 70.000 - 110.000
Meaningful equity
Competitive salary