Site Reliability Engineer

Reysion Technologies

Dresden

Vor Ort

EUR 90.000 - 130.000

Vollzeit

Vor 13 Tagen
Bewerbungsgenerator

Erhalte eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

Reysion Technologies is seeking a Site Reliability Engineer to strengthen platform operations with a focus on observability, secure logging, and automation. You will manage Kubernetes and container orchestration, develop automation scripts, and deploy IaC.

The role involves 24x7 on-call rotations, incident management, and adherence to SOPs to improve system stability. The ideal candidate will have hands-on experience with Elasticsearch and/or OpenSearch, Prometheus, Grafana, and CI/CD tools like

Qualifikationen

  • Strong Linux and Kubernetes fundamentals.

Aufgaben

  • Platform Engineering & DevOps: Manage Kubernetes and container orchestration, including Helm chart configurations and CI/CD pipelines (Jenkins, ArgoCD). Develop automation scripts (Python, Bash, Go) and deploy Infrastructure-as-Code (IaC) solutions.
  • Observability, Monitoring & Visualisation: Maintain Prometheus solutions (scrape configurations, alert rules, PromQL queries), administer Thanos and Grafana.
  • Elastic Stack Operations & Log Management: Configure and optimise Elasticsearch clusters, Logstash pipelines, and Kibana dashboards for secure, scalable log processing.
  • Incident Response, Troubleshooting & Collaboration: Participate in 24x7 on-call rotations for rapid incident response, troubleshoot platform, data and performance issues, and engage in Major Incident Management (MIM).
  • Secure Operations & Compliance: Ensure system operations meet security and data protection requirements, maintain secure documentation, and manage access control policies.

Kenntnisse

Kubernetes
Linux
Python
Go
Bash
Git-based workflows
PromQL
ELK Stack
Grafana
CI/CD
Helm
Jenkins
ArgoCD
OpenSearch
Elasticsearch
Cloud platforms
Security/compliance

Tools

Elasticsearch
OpenSearch
Kibana
Logstash
Grafana
Jenkins
ArgoCD
Helm
Terraform

Jobbeschreibung

About the Role:

We are seeking a Site Reliability Engineer (SRE) with a strong background in observability, secure logging, and automation. The ideal candidate will have hands-on experience with Elasticsearch and/or Prometheus platforms. This role encompasses critical responsibilities in platform operations, including incident management, execution of scheduled maintenance, and contributing to engineering tasks focused on enhancing system stability. The SRE will also be responsible for adhering to standard operating procedures (SOPs) and actively contributing to their continuous improvement by providing constructive feedback.

Key Responsibilities:

  • Platform Engineering & DevOps: Manage Kubernetes and container orchestration, including Helm chart configurations and CI/CD pipelines (Jenkins, ArgoCD). Develop automation scripts (Python, Bash, Go) and deploy Infrastructure-as-Code (IaC) solutions.
  • Observability, Monitoring & Visualisation: Maintain Prometheus solutions (scrape configurations, alert rules, PromQL queries), administer Thanos and Grafana.
  • Elastic Stack Operations & Log Management: Configure and optimise Elasticsearch clusters, Logstash pipelines, and Kibana dashboards for secure, scalable log processing.
  • Incident Response, Troubleshooting & Collaboration: Participate in 24x7 on-call rotations for rapid incident response, troubleshoot platform, data and performance issues, and engage in Major Incident Management (MIM).
  • Secure Operations & Compliance: Ensure system operations meet security and data protection requirements, maintain secure documentation, and manage access control policies.

Qualifications, Requirements, and Skills:

  • Strong grasp of Linux concepts, preferably in Kubernetes environments.
  • Solid understanding of networking fundamentals and REST APIs.
  • Proficiency in Python, Go, or Bash.
  • Proficiency in Git-based configuration management workflows.
  • Strong experience with the ELK Stack, Prometheus, and Grafana.
  • Experience with Elasticsearch and/or OpenSearch.
  • Knowledge of Kubernetes and familiarity with cloud platforms.
  • Familiarity with CI/CD tools like Helm, Jenkins, or ArgoCD.
  • Good understanding of PromQL is an advantage.
  • Fluent English communication skills.
  • Willingness to work shift-based 24x7 on-call support, including weekends and holidays.
  • Must possess Ü2 security clearance.
  • Citizenship required: Member state of the EU and NATO. No dual citizenship outside these countries.
  • Must reside in Germany and hold a German labor contract.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer
Site Reliability Engineer

HCLTech Germany • Deutschland

Vor Ort
EUR 90.000 - 120.000
Sovereign Cloud Engineer (m/w/d)
Sovereign Cloud Engineer (m/w/d)

United States Digital Space LLC • Walldorf

Vor Ort
EUR 90.000 - 130.000
Remote work
Hybrid work
Site Reliability Engineer
Site Reliability Engineer

Apprize Technology Solutions • Deutschland

Vor Ort
EUR 70.000 - 90.000
Senior Site Reliability Engineer / Kubernetes
Senior Site Reliability Engineer / Kubernetes

Jobgether • Deutschland

Vor Ort
EUR 90.000 - 120.000
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder • Berlin

Vor Ort
EUR 110.000 - 150.000
Hybrid work model
Flexible work policy
AI-driven environment
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FACT-Finder • Pforzheim

Vor Ort
EUR 90.000 - 125.000
Hybrid work model
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE – Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Vor Ort
Confidential
Hybrid work model
Site Reliability Engineer (SRE) – Kubernetes/Platform - Berlin/Frankfurt - €110,000–120,000
Site Reliability Engineer (SRE) – Kubernetes/Platform - Berlin/Frankfurt - €110,000–120,000

Findr • Berlin

Vor Ort
EUR 110.000 - 120.000
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Vor Ort
EUR 90.000 - 140.000
Hybrid work model (3 office days/week)
Staff Site Reliability Engineer (f/m/d)
Staff Site Reliability Engineer (f/m/d)

IONOS • Berlin

Vor Ort
EUR 85.000 - 110.000
Canteen subsidy
Free drinks
Employee discounts
+3