Senior Network Site Reliability Engineer – NetSRE

Jobtailor

Deutschland

Vor Ort

EUR 85.000 - 110.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Jobtailor is seeking a reliability-focused engineer in Germany to own service SLIs/SLOs, incident response, and post‑mortems. You will drive observability, design safe change workflows, and partner with network engineers to embed operability into designs.

You should bring strong Linux fundamentals, networking knowledge, and hands-on HA experience, plus ability to automate with Go or Python and use modern IaC/CI‑CD tooling to streamline operations.

Qualifikationen

  • Strong production Linux fundamentals and structured debugging of complex systems.
  • Solid understanding of networking basics and failure modes.
  • Experience operating high-availability systems and driving improvements.
  • Ability to write and maintain software/automation (Go or Python).
  • Experience with modern infrastructure tooling (IaC, CI/CD, containers).

Aufgaben

  • Define and own reliability goals for network services and critical paths (SLIs/SLOs, availability targets, error budgets).
  • Drive reliability improvements across the whole network: services, site readiness, inter-site connectivity, and operational standards.
  • Own incident response for your areas, lead investigations/post-mortems, and implement durable fixes.
  • Build and evolve observability: metrics/logs/traces, alerts, and faster debug loops during incidents.
  • Design safer change workflows: automation, CI/CD, test/staging environments, canarying, rollbacks, and auditability for network changes.
  • Work closely with network engineers and platform teams to embed operability into designs.

Kenntnisse

Linux fundamentals
Networking basics
Incident response
Observability
CI/CD automation

Tools

Terraform
Kubernetes
CI/CD pipelines

Jobbeschreibung

Responsibilities
  • Define and own reliability goals for network services and critical paths (SLIs/SLOs, availability targets, error budgets where it makes sense)
  • Drive reliability improvements across the whole network: not only services, but also site readiness, inter‑site connectivity (DCI), and operational standards
  • Own incident response for your areas, lead investigations/post‑mortems, and turn failures into durable fixes (not repeated firefighting)
  • Build and evolve observability: actionable metrics/logs/traces, alerting, and faster debug loops during and after incidents
  • Design safer change workflows: automation, CI/CD, test/staging environments, canarying, rollbacks, and auditability for network changes
  • Work closely with network engineers and platform teams to embed operability into designs and keep operations practical and fast
Requirements
  • Strong production Linux fundamentals and a structured approach to debugging complex systems
  • Solid understanding of networking basics and how real networks fail (control plane vs data plane, latency/loss, failure domains, etc.)
  • Hands‑on experience operating high‑availability systems and improving them over time (not just "keeping lights on")
  • Ability to write and maintain software/automation (Go is common for us; Python is also welcome)
  • Experience with modern infrastructure tooling (e.g., IaC, CI/CD, container platforms) and comfort automating operational workflows
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer
Site Reliability Engineer

Apprize Technology Solutions • Deutschland

Vor Ort
EUR 70.000 - 90.000
Senior Site Reliability Engineer – Kubernetes Platform
Senior Site Reliability Engineer – Kubernetes Platform

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Stuttgart

Vor Ort
EUR 70.000 - 100.000
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Deutschland

Remote
EUR 75.000 - 110.000
Senior Site Reliability Engineer – Hardware Automation
Senior Site Reliability Engineer – Hardware Automation

Jobtailor • Deutschland

Remote
EUR 65.000 - 95.000
Senior Network Security Engineer, Automation, Python
Senior Network Security Engineer, Automation, Python

Jobtailor • Deutschland

Remote
EUR 90.000 - 120.000
Senior Network Security and Automation Engineer, Python
Senior Network Security and Automation Engineer, Python

Jobtailor • Deutschland

Remote
EUR 90.000 - 130.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Visa Hunt • Deutschland

Vor Ort
EUR 90.000 - 150.000
Home office budget
Learning & development budget of €1000
Competitive salary
+5
Network Operations Engineer
Network Operations Engineer

Signify Technology • München

Vor Ort
EUR 80.000 - 100.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Jobtailor • Deutschland

Vor Ort
EUR 60.000 - 80.000
Health benefits
Financial benefits
Family benefits
+2