Site Reliability Engineering Architect

Cavendish Professionals

Berlin

Hybrid

EUR 80.000 - 110.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die genau auf das zugeschnitten sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

A global consulting organization is seeking an experienced Site Reliability Engineering (SRE) Architect to enhance their engineering leadership team in Berlin. The role involves defining and implementing reliability and performance strategies for cloud-native systems. Candidates should have over 10 years in software engineering and a strong background in cloud platforms, Kubernetes, and observability. This is a full-time mid-senior level position with opportunities for mentoring and thought leadership.

Qualifikationen

  • 10+ years in software engineering, with 5+ years in senior SRE/architecture roles.
  • Expertise in at least one major cloud provider.
  • Strong hands-on experience with Kubernetes and microservices.

Aufgaben

  • Design highly scalable and fault-tolerant infrastructure.
  • Define and govern SLOs, SLIs, and error budgets.
  • Lead observability design for metrics and logging.

Kenntnisse

Cloud infrastructure design
Kubernetes expertise
Infrastructure as Code (Terraform)
Observability platforms
Python
Go
Distributed systems knowledge

Tools

AWS
GCP
Azure
Prometheus
Grafana
Ansible
Terraform

Jobbeschreibung

Overview

Site Reliability Engineering (SRE) Architect – Munich (Remote/Hybrid)

I am partnered with a global consulting organisation that is expanding its engineering leadership team in Germany. They are seeking an experienced SRE Architect to define and drive the long-term reliability, scalability, and performance strategy across complex, cloud-native systems. This is a senior architectural role with wide influence across engineering, combining deep technical expertise with leadership. You will set standards, frameworks, and practices that enable teams to deliver world-class services at scale.

Key Responsibilities
  • Architect & Strategy – Design highly scalable and fault-tolerant infrastructure on leading cloud platforms (AWS, GCP, or Azure).
  • Reliability Frameworks – Define and govern SLOs, SLIs, and error budgets across engineering teams.
  • Observability – Lead observability design for metrics, tracing, logging, and alerting.
  • Automation & IaC – Champion Infrastructure as Code (Terraform, Ansible) for secure, repeatable provisioning.
  • Resilience & Recovery – Develop disaster recovery strategies, resilience patterns, and chaos engineering practices.
  • Leadership & Mentoring – Act as a thought leader, mentoring engineers and embedding reliability best practices across the organisation.
  • Incident Evolution – Analyse major incidents, drive systemic improvements, and evolve incident management culture.
Key Requirements
  • 10+ years in software engineering, DevOps, or systems engineering, including 5+ years in senior SRE/architecture roles.
  • Expertise in at least one major cloud provider (AWS, GCP, Azure).
  • Strong hands-on experience with Kubernetes and microservices at scale.
  • Proven skills in Infrastructure as Code (Terraform, Ansible, Chef, or Puppet).
  • Solid background in observability platforms (Prometheus, Grafana, OpenTelemetry, ELK, Datadog, etc.).
  • Proficiency in Python or Go for automation and tooling.
  • Deep knowledge of distributed systems, networking, and high-availability design patterns.
Nice-to-Haves
  • Professional cloud certifications.
  • Knowledge of service mesh technologies
  • DevSecOps/security best practices.
  • Experience leading large-scale tech transformations.

If this sounds like something you’d thrive in—or even if you’re just curious—I’d love to chat and tell you more.

Cavendish (Recruitment) Professionals Ltd are proud to be an equal opportunity employer and we believe that inclusivity begins with the candidate experience. All qualified applicants will receive consideration for employment regardless of gender, race, age, sexual orientation, religion, or belief.

Seniority level
  • Mid-Senior level
Employment type
  • Full-time
Job function
  • Information Technology
Industries
  • IT Services and IT Consulting
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer / Kubernetes
Senior Site Reliability Engineer / Kubernetes

Jobgether • Deutschland

Vor Ort
EUR 90.000 - 120.000
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Vor Ort
Confidential
Hybrid work model
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder • Berlin

Vor Ort
EUR 110.000 - 150.000
Hybrid work model
Flexible work policy
AI-driven environment
Site Reliability Engineer (SRE) – Kubernetes/Platform - Berlin/Frankfurt - €110,000–120,000
Site Reliability Engineer (SRE) – Kubernetes/Platform - Berlin/Frankfurt - €110,000–120,000

Findr • Berlin

Vor Ort
EUR 110.000 - 120.000
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder • Pforzheim

Vor Ort
EUR 90.000 - 125.000
Hybrid work model
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Meyandy LLC • Berlin

Vor Ort
EUR 90.000 - 130.000
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Vor Ort
EUR 90.000 - 140.000
Hybrid work model (3 office days/week)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)

Fact Finder • Berlin

Vor Ort
EUR 90.000 - 120.000
Site Reliability Engineer
Site Reliability Engineer

Reysion Technologies • Dresden

Vor Ort
EUR 90.000 - 130.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Ageras Danmark • Berlin

Vor Ort
EUR 70.000 - 90.000
Impactful role
Growth opportunities
Collaborative team
+1