Sovereign Cloud Engineer (m/f/d)

VASPP

Deutschland

Vor Ort

EUR 90.000 - 130.000

Vollzeit

Vor 2 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Eine komplette Bewerbung in einer Minute — maßgeschneiderter Lebenslauf und Anschreiben, versandbereit.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

VASPP in Germany seeks an experienced Sovereign Cloud / Site Reliability Engineer (SRE) to support secure, reliable, and continuous operation of cloud-native Kubernetes platforms.

You will join an experienced DevOps/SRE team, owning platform operations, automation, monitoring, incident management, security, and continuous improvement for a highly available Unified Observability Platform in a 24x7 environment.

Qualifikationen

  • Experience with Kubernetes and cloud platforms.
  • Strong knowledge of CI/CD and IaC practices.
  • Security-conscious with 24/7 operations.
  • Proficiency in Python/Go scripting.

Aufgaben

  • Operate and maintain Kubernetes clusters with Gardener.
  • Triage incidents and ensure high availability.
  • Optimize CI/CD pipelines with Jenkins and ArgoCD.
  • Develop automation tooling using Python/Go.
  • Monitor observability with Prometheus and OpenTelemetry.

Kenntnisse

Kubernetes operations
Incident management
Automation scripting
Security mindset

Tools

Gardener
Helm
Jenkins
ArgoCD
Git
IaC
Python
Go
Bash
Prometheus
OpenTelemetry
Grafana
Elasticsearch/OpenSearch
Logstash
Kibana
ServiceNow
PagerDuty

Jobbeschreibung

Before You Apply – Mandatory Requirements

Due to the security requirements of the role, candidates must meet all of the following:

  • Citizenship: Citizenship of a country that is a full member of both the EU and NATO. If you hold multiple citizenships, all must meet this requirement.
  • Residence & Employment: Residence in Germany and direct employment through a German legal entity under a German employment contract and German labor law. Employment through foreign entities or subcontractors is not permitted.
  • Security Clearance: A valid and verifiable Ü2 security clearance in accordance with the German Security Clearance Act (SÜG) is mandatory.
Overview

For our growing team, we are looking for an experienced Sovereign Cloud / Site Reliability Engineer (SRE) to support the secure, reliable, and continuous operation of modern cloud-native and Kubernetes platforms.

You will work with an experienced DevOps/SRE team on highly secure, business-critical platforms, taking responsibility for platform operations, automation, monitoring, incident management, security, and continuous improvement.

The role focuses particularly on Kubernetes, CI/CD, Infrastructure as Code, observability, logging, and operational excellence within a highly available Unified Observability Platform in a 24x7 operational environment.

Key Technologies

Kubernetes | Gardener | Helm | Jenkins | ArgoCD | Git | IaC | Python | Go | Bash | Prometheus | PromQL | Thanos | OpenTelemetry | Grafana | Elasticsearch | OpenSearch | Logstash | Kibana | ServiceNow | PagerDuty | REST APIs

Key Responsibilities
Kubernetes & Platform Operations
  • Operate and maintain Kubernetes clusters with Gardener, including workloads, deployments, Helm charts and platform components.
  • Troubleshoot availability, performance and deployment issues.
  • Ensure secure, scalable, resilient and highly available platforms.
CI/CD & Automation
  • Operate and optimize Jenkins and ArgoCD pipelines for automated deployments.
  • Implement IaC and Git-based deployment workflows.
  • Develop automation and operational tools using Python, Go and/or Bash.
  • Automate provisioning, health/compliance checks, alerting and reporting.
Monitoring & Observability
  • Manage Prometheus, Thanos and OpenTelemetry environments, including scrape jobs, alert rules and PromQL.
  • Develop and maintain Grafana dashboards and continuously improve monitoring and alerting.
Logging & Log Management
  • Operate and optimize Elasticsearch/OpenSearch, Logstash and Kibana.
  • Monitor log ingestion, storage, performance and reliability.
  • Support centralized troubleshooting, anomaly detection, security and compliance.
Integration & Operations
  • Integrate observability and logging platforms with ServiceNow, PagerDuty and other enterprise tools via secure APIs.
  • Handle operational requests and participate in Scrum, DevOps and service-improvement activities.
  • Collaborate with internal teams, SAP, suppliers and stakeholders.
Incident & Problem Management
  • Participate in a 24/7 on-call and shift rotation, including weekends and public holidays.
  • Respond to platform, monitoring, logging and deployment incidents.
  • Perform RCA, support Major Incident Management (MIM) and implement sustainable corrective actions.
  • Continuously improve platform stability, resilience and operational processes.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Sovereign Cloud Engineer (m/w/d)
Sovereign Cloud Engineer (m/w/d)

United States Digital Space LLC • Walldorf

Hybrid
EUR 90.000 - 130.000
Remote work
Hybrid work
Sovereign Cloud Engineer (m/w/d)
Sovereign Cloud Engineer (m/w/d)

TMT Prüfservice GmbH & Co KG • Walldorf

Hybrid
Confidential
Site Reliability Engineer
Site Reliability Engineer

Roc Search • Deutschland

Vor Ort
EUR 70.000 - 110.000
Network & Security Expert
Network & Security Expert

Mindquest • Berlin

Hybrid
EUR 65.000 - 90.000
Site Reliability Engineer - Kubernetes / DevOps (m/w/d)
Site Reliability Engineer - Kubernetes / DevOps (m/w/d)

StudentJob • Karlsruhe

Hybrid
EUR 60.000 - 90.000
Hybrid-Arbeitsmodell
Mitarbeitervergünstigungen
Mitarbeiterveranstaltungen
+2
Senior Site Reliability Engineer / Kubernetes
Senior Site Reliability Engineer / Kubernetes

Jobgether • Deutschland

Remote
EUR 90.000 - 120.000
Customer Operations Engineer 24/7(f/m/d)
Customer Operations Engineer 24/7(f/m/d)

1&1 IONOS SE • Berlin

Hybrid
EUR 70.000 - 100.000
Hybrid work
Modern office
Good transport
+3
Infrastructure, DevOps Architect (m/f/d)
Infrastructure, DevOps Architect (m/f/d)

Giesecke+Devrient GmbH • München

Vor Ort
EUR 150.000 - 210.000
Customer Operations Engineer 24/7(f/m/d)
Customer Operations Engineer 24/7(f/m/d)

IONOS Group • Berlin

Hybrid
EUR 68.000 - 100.000
Hybrid working model
Modern office with good transport
Employee discounts
+2
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)
Senior Site Reliability Engineer / SRE - Kubernetes & Hybrid Cloud (m/f/d)

FactFinder • Berlin

Vor Ort
EUR 90.000 - 140.000
Hybrid work model (3 office days/week)