Senior Platform Engineer (f/m/x)

Personio

Berlin

Hybrid

EUR 110.000 - 140.000

Vollzeit

Vor 6 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Eine maßgeschneiderte Bewerbung für diese Stelle — ein maßgeschneiderter Lebenslauf und ein Anschreiben, die genau zur Stellenanzeige passen.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Office, Remote & Workation
Flexible working hours
Excellent hardware provided
Team events
Education budget

Zusammenfassung

Workist is seeking a Senior Platform Engineer to own the AI agent platform. You will manage two clouds, Kubernetes clusters for production, staging and ML workloads, and drive the path from merge request to production.

You set the roadmap and ensure security, observability and cost-efficiency. You will design platform capabilities, including GPU capacity planning and an LLM gateway with region failover, while aligning with quarterly themes and incident response.

Qualifikationen

  • Proven production infrastructure ownership and end-to-end responsibility.
  • Kubernetes and Terraform in production: upgrades, Helm, GitOps, IaC.
  • Production Azure or AWS experience; Azure dominant.

Aufgaben

  • Own the architecture of cloud environments across Azure and AWS.
  • Design and deliver platform capabilities: GPU capacity and LLM gateway.
  • Set the standard for Kubernetes and managed data services, with cost control.
  • Own security posture end to end: least-privilege, network isolation, patching.
  • Own observability and alerting; lead incident response and post mortems.
  • Run quarterly backup/restore and disaster-recovery tests.
  • Own the path from merge request to production: fast builds and reliable deployments.

Kenntnisse

Kubernetes
Terraform
Python
GitOps
CI/CD
Azure
AWS
Security
Observability
Cost optimization

Tools

Azure AKS
Helm
GitLab
Terraform
OpenSearch
Postgres
Redis
Kubernetes Gateway API

Jobbeschreibung

Your role with us

AsSenior Platform Engineer, you own the platform that Workist's AI agents run on: two clouds, the Kubernetes clusters for production, staging and ML workloads, and the path from merge request to production for every service we ship. You decide how this platform evolves, you keep it secure, observable and cost-efficient, and you extend it as the product grows. Most recently that meant an LLM gateway with region failover in front of our Azure OpenAI deployments, and performing GPU capacity planning for our own models.

You set your own roadmap. Infrastructure at Workist is run as quarterly themes that you propose, make the case for and deliver, and roughly a third of your time goes to whatever the week brings: an incident, a pentest finding, a developer whose deployment is stuck and needs a second pair of eyes. You report to our CTO and are the voice of the platform in engineering decisions.

How we build and run things
  • Everything runs on Kubernetes (Azure AKS), deployed with Helm and GitOps/Flux. All infrastructure is Terraform, applied through CI.

  • CI/CD for every service on GitLab

  • Managed services wherever they keep life simple: Postgres, OpenSearch, Redis, blob storage. We self-host only where it clearly pays off.

  • Two clouds: Azure as our primary cloud, AWS for search and mail ingestion.

  • Python everywhere: read and fix application code when that is where the fix belongs.

  • Security is routine: automated scanning in every pipeline, regular external pentests, quarterly backup and disaster-recovery tests.

Your responsibilities
Own the platform
  • Own the architecture of our cloud environments across Azure and AWS: how subscriptions, networks, identities and environments are structured and stay isolated

  • Design and deliver the platform capabilities the product needs next, from GPU capacity and an LLM gateway to new environments

  • Set the standard for how we run Kubernetes and our managed data services: capacity planned ahead of growth, changes and upgrades rolled out safely and routinely, costs kept in check

Keep it quiet
  • Own the security posture end to end: least-privilege access, network isolation, findings from scanners and pentests, and a patch cadence that runs itself.Be a technical counterpart for audits and security questionnaires

  • Own observability and alerting so that every alert is worth acting on, and lead the response to platform incidents including the follow-up fixes

  • Run quarterly backup/restore and disaster-recovery tests, and close what they reveal

Make the team faster
  • Own the path from merge request to production: fast builds, reliable deployments, environments and access on demand

  • Build the infrastructure behind our AI models: the LLM gateway, model deployments with region failover, GPU capacity, and whatever the next model needs

What this looked like recently: a Redis instance that silently dropped half of its new TCP connections and took the Celery workers down with it, region failover and load balancing for our Azure OpenAI deployments, moving our clusters from nginx ingress to the Kubernetes Gateway API.
Your profile

We're looking for a professional team player who thrives in an environment where we share knowledge openly and push each other to deliver the best possible outcome.What you bring:

  • You have run production infrastructure for years, ideally as the person who owned it end to end rather than one of many in a large ops team

  • Kubernetes and Terraform in production: cluster upgrades, Helm, GitOps, and infrastructure as code for a whole cloud estate

  • Production experience on Azure or AWS. Most of our estate is Azure, so an AWS background works if you are willing to go deep on Azure.

  • Active Python skills: you read and debug a Python codebase and fix the application when the application is what is broken

  • Security as a habit: you have run patch management, handled pentest findings and designed least-privilege access, and you prioritise findings pragmatically

  • Youwrite clearly: runbooks, incident write-ups and architecture decisions that others can follow

  • Fluent English(engineering runs in English, so German is not required, though it can help in some customer and vendor conversations)

Our benefits & culture
  • Worki Tower: Our office in the heart of the city over 8 floors invites you to work professionally at an ergonomic workplace and gives you space for your creativity. Our own Workist Lounge awaits you with a cozy atmosphere including coffee, drinks and a vitamin basket for short or longer creative breaks. Play table tennis, darts or Nintendo to clear your head and exchange ideas with other teams.
  • Flexible working hours: Shape your working day with our flexible working hours. Have important appointments in the morning or evening? No problem, take advantage of work schedule flexibility.
  • Office, Remote & Workation: You decide yourself and have the flexible choice between an office or remote from your favorite location! We give you the freedom to work 4 weeks a year from abroad. Your state-of-the-art hardware will accompany you on your way!
  • Come-Together: Let's play together - let's celebrate together! End the working day with a cool drink in our lounge or enjoy the time with your colleagues at the summer party, Christmas party or our quarterly team events.
  • Stay fit & healthy: Your fitness is important to us, so join our weekly RunningLunch and connect with your colleagues. Work-life balance and reduction of your stress level included.
  • Stay mobile:Choose between Deutschlandticket Job and a JobRad allowance and be mobile throughout Berlin and/or Germany.
  • Be social: For your heart's project or to get involved in social and volunteer work, we support you with 3 additional "Social Days".
  • Further your education - with a personal development budget of EUR 1,000.00 per year.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Platform Engineer (f/m/x) arbeitnow Workist Gmbh Berlin · 9/27/2026
Senior Platform Engineer (f/m/x) arbeitnow Workist Gmbh Berlin · 9/27/2026

Primetime • Berlin

Hybrid
EUR 110.000 - 150.000
Office & remote flexibility
Flexible working hours
Educational budget EUR 1,000 per year
+1
Senior Platform Engineer (f/m/x)
Senior Platform Engineer (f/m/x)

Workist Gmbh • Berlin

Vor Ort
EUR 90.000 - 130.000
Office in city center
Table tennis & darts
Flexible working hours
+1
Senior Platform Engineer (f/m/x)
Senior Platform Engineer (f/m/x)

Workist • Berlin

Hybrid
EUR 120.000 - 180.000
Office & remote
Flexible hours
Workation
+5
Cloud Platform Engineer — Kubernetes, CI/CD & APIs
Cloud Platform Engineer — Kubernetes, CI/CD & APIs

Polarise SARL • Deutschland

Remote
USD 104.000 - 162.000
Remote work model
Welcome package with company merch
MacBook Pro or ThinkPad hardware
Senior DevOps Engineer
Senior DevOps Engineer

HR New Media GmbH • Berlin

Hybrid
EUR 90.000 - 130.000
Hybrid workstyle
30 days annual vacation
Two volunteering days
Platform Engineer (m/f/d)
Platform Engineer (m/f/d)

Redcare Pharmacy • Köln

Vor Ort
EUR 80.000 - 120.000
Welcome days
Sports membership
Mental health support
+2
Senior DevOps / Platform Engineer (F/M/X)
Senior DevOps / Platform Engineer (F/M/X)

Remotely • Berlin

Vor Ort
EUR 90.000 - 130.000
30 vacation days
Training budget €1500/year
BVG ticket
+3
(Senior) DevOps Engineer (m/f/x)
(Senior) DevOps Engineer (m/f/x)

AERTiCKET AG  • Berlin

Hybrid
EUR 85.000 - 115.000
Hybrid work model
Office amenities
Training budget
+1
Head of AI Engineering (f/m/x)
Head of AI Engineering (f/m/x)

Neoshare • Berlin

Vor Ort
EUR 120.000 - 180.000
30 vacation days
Jobticket
Urban Sports/EGYM subsidy
+2
Senior Software Engineer in Test (f/m/d/x)
Senior Software Engineer in Test (f/m/d/x)

DUDE CHEM • Berlin

Hybrid
EUR 85.000 - 120.000