Senior Platform & Reliability Engineer (all genders)

Meyandy LLC

Berlin, Hamburg, Köln, München

Hybrid

EUR 110.000 - 150.000

Vollzeit

Vor 3 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Hebe dich für diese Rolle ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Contabo is seeking a Senior Platform & Reliability Engineer. This remote-first role offers the choice of hybrid or on-site work in Germany. You will own on-call processes, advance observability across the stack, and advise multiple teams on shared infrastructure services.

Strong SRE and storage experience are essential. Key focus areas include Kong/Nginx ingress, Ceph/Longhorn storage, Vault/Keycloak identity, Cloudflare WAF, and distributed tracing with OpenTelemetry.

Qualifikationen

  • 7+ years of experience in platform, infrastructure, or SRE roles with end-to-end responsibility.
  • Hands-on with Ceph and Kubernetes persistent storage.
  • Experience operating API gateways and ingress (Kong, Nginx Ingress).
  • Experience with CDN/edge security, DDoS mitigation, and WAF configuration (Cloudflare).
  • Strong grounding in system design fundamentals (load balancing, caching, sharding, replication).
  • Experience designing or maturing observability with Prometheus, Grafana, and OpenTelemetry.
  • Experience with Vault and Keycloak; German language skills are a plus.

Aufgaben

  • Design and establish an on-call process with runbooks for platform and infra incidents.
  • Mature observability practices, roll out tracing, SLOs/SLIs, and dashboards.
  • Provide technical authority across teams for shared infrastructure services.
  • Lead system design reviews and apply fundamentals to new and existing services.
  • Document knowledge and prevent single points of failure across platform-wide systems.
  • Occasional travel for datacenter visits and team offsites; remote-first with hybrid/on-site options.

Kenntnisse

Platform & SRE
Ceph storage
Longhorn storage
Kubernetes storage
API gateways & ingress
Cloudflare WAF
Observability
On-call processes
Documentation
English fluency

Tools

Prometheus
Grafana
OpenTelemetry
Vault
Keycloak
NATS
Proxmox
OpenStack

Jobbeschreibung

**Your creative field**We are looking for a full-time, permanent Senior Platform & Reliability Engineer (all genders) to start as soon as possible. We live remote-first, but you have the freedom to choose whether you want to work hybrid or completely on-site due to your proximity to one of our locations (Berlin, Cologne, Hamburg, Munich). As a Senior Platform & Reliability Engineer at Contabo, you take architectural ownership of our shared infrastructure services – the foundation that multiple development teams build on: API gateway and ingress (Kong, Nginx Ingress), persistent storage (Ceph, Longhorn), secrets and identity infrastructure (Vault, Keycloak), edge security and WAF (Cloudflare), and our observability stack. You'll be taking over grown, partially under-documented systems – and that's exactly where the appeal of this role lies: you work your way deep into these systems, identify and remediate known weak points, evaluate aging components with solution-agnostic build-vs-buy reasoning, and decide which legacy pieces get fixed, replaced, or retired. One of your central mandates: you design and establish an on-call process, including runbooks for platform and infrastructure incidents – where no formal process exists today. In parallel, you mature our observability practices, rolling out distributed tracing, SLOs/SLIs, and meaningful dashboards across the stack, building on our existing tooling with Prometheus, Grafana, Alloy, and OpenTelemetry. In system design reviews, you bring your strong grounding in the fundamentals – load balancing, caching, sharding, replication, consistency trade-offs – and apply them to new and existing services alike. You won't be managing a team, but you will be the technical authority multiple teams rely on: you advise across teams on shared infrastructure services, document your knowledge consistently, and actively distribute it – so that critical know-how never again depends on a single person. Success in this role means: known risks are resolved, a solid on-call process is up and running, and single points of failure are measurably reduced platform-wide. The position is remote (Germany), with hybrid or on-site work optional; occasional travel for datacenter visits and team offsites is part of the role. **What convinces us**Your personality, paired with: * 7+ years of experience in platform, infrastructure, or SRE roles, ideally with end-to-end responsibility for a private cloud or IaaS platform* Hands-on production experience with distributed storage systems (Ceph) and Kubernetes persistent storage (Longhorn or comparable)* Experience operating API gateways and ingress (Kong, Nginx Ingress, or comparable), including debugging cross-cutting concerns like CORS and rate limiting* Experience with CDN/edge security, DDoS mitigation, and WAF configuration (e.g. Cloudflare), as well as firewall-rule design* A strong grounding in system design fundamentals (load balancing, caching, sharding/replication, consistency models, message queues) – and the judgment to apply them to real-world trade-offs* Experience designing or maturing observability (tracing, metrics, SLOs/SLIs) with tools such as Prometheus, Grafana, and OpenTelemetry is desirable* Experience with secrets/identity infrastructure (Vault, Keycloak), messaging systems (NATS), and building on-call processes and incident runbooks from scratch is a plus* Composure working with grown, incompletely documented systems – paired with the right mix of pragmatism and perfectionism: fix what's broken, retire what's dead, rather than rewriting everything* Genuine enjoyment of acting as the technical go-to person across teams, actively sharing and documenting your knowledge* Professional fluency in English (our working language); German language skills, certifications (CKA/CKS, Ceph training), and experience with virtualization platforms (Proxmox, OpenStack) are a plus **Awesome Prospects**At Contabo, we are constantly evolving. Our growth creates space for new ideas, ownership, and real impact for people who want to make a difference. What defines us is trust, direct communication, and an open feedback culture. We believe the best solutions are built togethe...
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Platform & Reliability Engineer (all genders)
Senior Platform & Reliability Engineer (all genders)

Contabo GmbH • München

Hybrid
EUR 90.000 - 120.000
Work the way you live - remote/hybrid
EU workation & Mallorca office
Wellpass fitness program
+6
Senior Platform & Reliability Engineer (all genders)(EN)
Senior Platform & Reliability Engineer (all genders)(EN)

Contabo • Deutschland

Hybrid
EUR 90.000 - 130.000
Remote or hybrid options
Flexible hours
Travel for datacenters/team offsites
+5
Senior Software Developer – Platform & Infrastructure Engineering (all genders)
Senior Software Developer – Platform & Infrastructure Engineering (all genders)

Meyandy LLC • Deutschland

Hybrid
EUR 90.000 - 130.000
Remote-friendly in Germany
Hybrid or on-site options
EU work location flexibility
+2
Senior Software Developer – Platform & Infrastructure Engineering (all genders)
Senior Software Developer – Platform & Infrastructure Engineering (all genders)

Contabo • Deutschland

Vor Ort
EUR 90.000 - 130.000
Remote-friendly
Hybrid options
Flexible hours
+3
Senior Software Developer – Platform & Infrastructure Engineering (all genders)
Senior Software Developer – Platform & Infrastructure Engineering (all genders)

Contabo GmbH • Deutschland

Remote
EUR 90.000 - 120.000
Remote or hybrid work optional
Workation across EU
EGYM Wellpass & wellness facilities
+7
Senior Software Developer – Platform & Infrastructure Engineering (all genders)
Senior Software Developer – Platform & Infrastructure Engineering (all genders)

Contabo GmbH • München

Hybrid
EUR 90.000 - 120.000
Remote or hybrid
EU Workation
Wellness benefits
+7
Senior Software Developer – Platform & Infrastructure Engineering (all genders)(EN)
Senior Software Developer – Platform & Infrastructure Engineering (all genders)(EN)

Contabo • Deutschland

Hybrid
EUR 90.000 - 130.000
Remote work
Hybrid option
EGYM Wellpass
+5
Senior Frontend Developer (all genders)
Senior Frontend Developer (all genders)

Meyandy LLC • Deutschland

Hybrid
EUR 70.000 - 100.000
Remote or hybrid work
Flexible hours
Annual vacation
Senior Cloud Product Manager (all genders)
Senior Cloud Product Manager (all genders)

Contabo GmbH • Deutschland

Remote
EUR 95.000 - 150.000
Remote or hybrid with flexible hours
Workation across the EU and Mallorca
EGYM Wellpass access
+1
Senior Network Engineer (all genders)
Senior Network Engineer (all genders)

Contabo GmbH • München

Hybrid
EUR 85.000 - 120.000
Remote or hybrid with flexible hours
EU workation opportunities
EGYM Wellpass access
+2