Senior Site Reliability Engineer

Visa Hunt

Deutschland

Vor Ort

EUR 90.000 - 150.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Home office budget
Learning & development budget of €1000
Competitive salary
Variable remuneration program
Monthly meal allowance
Deutschland ticket subsidy
28 vacation days
Opportunity to work abroad for up to

Zusammenfassung

Visa Hunt is seeking a senior DevOps/SRE engineer to design monitoring standards and resilient systems. You will implement tooling to automate operations, build dashboards, and translate high-level goals into actionable tasks.

Ideal candidates have 6+ years in relevant fields, strong programming skills (Python/Go/Java/Ruby), and deep experience with service meshes and chaos testing. German is a plus; English must be strong.

Qualifikationen

  • Degree in Computer Science, Software Engineering, Information Technology or equivalent professional experience.
  • 6+ years in DevOps, SRE, or Software Engineering roles in a high-growth environment.
  • Programming experience in Python, Ruby, Java, or Go.
  • Expertise in Service Discovery / Service Mesh to manage microservice communications.
  • Experience in capacity analysis and preemptive addressing of SLO exhaustion.
  • Designing systems with cross-DC failover capabilities.
  • Experience with Chaos Mesh or Gremlin for proactive testing.
  • Designing infrastructure to minimize blast radius of failures.
  • Participation in a 24/7 on-call rotation.
  • Business-proficient English; German a plus.

Aufgaben

  • Design monitoring and alerting standards, including SLIs, SLOs, and SLAs.
  • Implement software and tooling to improve resilience and automate operations.
  • Create dashboards for generic services and product-specific monitoring, breaking goals into deliverables.

Kenntnisse

DevOps
SRE
Software Engineering
Python
Ruby
Java
Go
Service Mesh
Chaos engineering
On-call rotation

Ausbildung

Bachelor's degree in CS/Software Engineering

Tools

Chaos Mesh
Gremlin

Jobbeschreibung

Role Overview

Responsible for the design of monitoring and alerting standards, creating frameworks that implement key concepts like SLIs, SLOs, and SLAs, and championing standards for incident response and reliability best practices.

What You Will Do

Implement software and tooling to improve resilience and automate operations, create generic service and product specific monitoring dashboards, and break down high-level architectural goals into small, deliverable tasks.

Why It Might Be a Fit

We're looking for a curious self-starter who combines openness and creativity with a structured, hands‑on mentality, and brings an agile mindset along with the drive to get things done and the self‑reflection to keep improving.

Requirements
  • Degree in Computer Science, Software Engineering, Information Technology or equivalent professional experience
  • 6+ years in DevOps, SRE, or Software Engineering roles in a high‑growth environment
  • Programming experience in Python, Ruby, Java, or Go
  • Expertise in Service Discovery / Service Mesh to manage microservice communications
  • Experience in the analysis of capacity requirements and address exhaustion before it impacts SLOs
  • Designing systems that automatically failover across different data centers
  • Expertise in tools like Chaos Mesh or Gremlin to proactively test system weaknesses
  • Designing infrastructure with a focus to limit the blast radius of any single failure
  • Experience in the participation of a 24/7 on‑call rotation
  • Business proficient English (written and verbal). German language is a plus.
Benefits
  • Home office budget
  • Learning & development budget of €1000 per year
  • Competitive salary
  • Variable remuneration program
  • Monthly meal allowance
  • Deutschland ticket subsidy
  • 28 vacation days
  • Opportunity to work abroad for up to 12 weeks per year
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobgether • Deutschland

Vor Ort
EUR 90.000 - 130.000
Fully remote
High ownership
Tech exposure
Site Reliability Engineer (m/f/d)
Site Reliability Engineer (m/f/d)

Solactive AG • Berlin

Vor Ort
Vertraulich
30 annual vacation days
Job ticket
Gym access
+2
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Stuttgart

Vor Ort
EUR 70.000 - 100.000
Site Reliability Engineer (m/w/d)
Site Reliability Engineer (m/w/d)

PAYBACK • München

Vor Ort
EUR 55.000 - 75.000
Delicious meals in canteen
24/7 access to gym
Flexible working hours
+3
(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz
(Senior) Site Reliability Engineer (m/f/d) in Berlin or Konstanz

United States Digital Space LLC • Deutschland

Hybrid
EUR 60.000 - 90.000
Hybrid working arrangements
Flexible hours
Subsidized sports or yoga courses
+1
Site Reliability Engineering Architect
Site Reliability Engineering Architect

Cavendish Professionals • Berlin

Hybrid
EUR 80.000 - 110.000
Senior Site Reliability Engineer (all genders)
Senior Site Reliability Engineer (all genders)

FACT-Finder • Pforzheim

Hybrid
EUR 90.000 - 140.000
Hybrid work model
Impact on product reliability
Competitive compensation
+1
Senior Site Reliability Engineer (m/f/d)
Senior Site Reliability Engineer (m/f/d)

TOPdesk • Kaiserslautern

Vor Ort
EUR 90.000 - 150.000
30 days annual vacation
Remote-friendly options
Health and wellness programs
+2
Senior Site Reliability Engineer (m/w/d)
Senior Site Reliability Engineer (m/w/d)

Impower • München

Hybrid
EUR 70.000 - 90.000
Flexible hours
Ownership in projects
Diverse team culture
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Solaris SE • Berlin

Vor Ort
EUR 75.000 - 95.000
Home office budget
Learning budget €1000/yr
Monthly meal allowance
+2