Senior SRE: Kubernetes Reliability & GitOps Lead

SysEleven GmbH

Germany (OH)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A cloud services provider is seeking a Senior Site Reliability Engineer to design, build, and operate APIs for their as-a-Service products. The ideal candidate will have several years of experience in Kubernetes and Linux environments, strong problem-solving skills, and proficiency in Go. Responsibilities include ensuring the reliability of services, managing container applications, and leading incident response efforts. This role offers autonomy in driving initiatives and fostering a culture of knowledge sharing within the team.

Qualifications

  • Several years of experience in Linux and Kubernetes environments.
  • Strong understanding of observability concepts.
  • Practical development experience in Go.

Responsibilities

  • Ensure reliability and performance of Database and Observability as a Service.
  • Manage container-based applications in Kubernetes.
  • Lead incident response and root cause analysis.
  • Develop API services and tooling in Go.

Skills

Problem-solving skills
Communication skills in German
Communication skills in English

Tools

Go
Terraform
Kubernetes
GitLab CI
Argo CD
Python
Bash

Job description

Your mission

As a Senior Site Reliability Engineer (m/f/x) at SysEleven, you design, build, and operate APIs that power the automation and reliability of our as-a-Service products, such as Database as a Service. You use Infrastructure as Code to standardize and scale our platforms, and you continuously improve CI/CD pipelines to ensure secure, resilient, and efficient delivery processes. With GitOps practices and Kubernetes orchestration, you reduce operational complexity and enable stable, predictable deployments that support our customers’ critical workloads. You take ownership of reliability end to end, contribute to a culture of continuous improvement, and lead by example in solving complex technical challenges that shape the future of our services.

Your tasks

  • Ensure the reliability, availability, and performance of our Database- and Observability-as-a-Service products
  • Manage container-based applications in Kubernetes with a strong focus on security and resilience
  • Lead incident response, root cause analysis, and sustainable remediation efforts
  • Apply GitOps principles using Helm and Argo CD
  • Develop API services and tooling in Go to deliver stable SaaS products
  • Build and optimize CI/CD pipelines to improve deployment safety and system stability
  • Design and manage scalable infrastructure using IaC tools (e.g., Terraform) in cloud environments

Our Technologies And Tech Stack

  • Go, Python, Bash
  • OpenStack, Kubernetes, Cilium, Envoy, Kyverno
  • Terraform, Crossplane, Argo CD, GitLab CI
  • PostgreSQL, Grafana, Loki, Mimir

Requirements

  • Several years of experience operating highly available systems in Linux and Kubernetes environments
  • Strong understanding of observability concepts (monitoring, logging, tracing)
  • Practical development experience in Go (knowledge of Python or Rust is a plus)
  • Experience with Infrastructure-as-Code tools such as Terraform or OpenTofu
  • Hands-on experience in incident management and structured root cause analysis
  • Familiarity with CI systems, especially GitLab CI
  • Strong problem-solving skills and good communication skills in German and English (minimum B2 level)

What you can expect

At SysEleven, you take ownership of the reliability of customer-facing services such as Database as a Service and Observability as a Service, which are deeply integrated into our cloud and Kubernetes platforms.

You actively contribute to the daily operations and continuous improvement of these services, focusing on stability, performance, and automation maturity.

We value a blameless culture, open communication, and knowledge sharing — whether in day-to-day collaboration, internal “Show & Tell” sessions, or at external conferences. You will have the autonomy to drive reliability initiatives strategically and shape robust, sustainable platform solutions together with the team.

Contact

About Us

We are your partner for managed cloud and Kubernetes services - Made in Germany!

We take responsibility and stand for security, reliability, and scalability in the operation of your business-critical applications in Germany. We provide you with a secure cloud and network infrastructure - made in Germany, consulting and efficient Kubernetes operating models.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (m/f/x)
Senior Site Reliability Engineer (m/f/x)

SysEleven GmbH • Germany (OH)

On-site
USD 100,000 - 130,000
Senior Site Reliability Engineer (m/w/d)
Senior Site Reliability Engineer (m/w/d)

Impower • Germany (OH)

Hybrid
USD 81,000 - 105,000
Flexible working hours
Real ownership in projects
Growth opportunities in a modern environment
Senior Site Reliability Engineer (all genders)
Senior Site Reliability Engineer (all genders)

FACT-Finder • Germany (OH)

Hybrid
USD 103,000 - 139,000
Hybrid work model
Open feedback culture
Reliability discipline
(Senior) Site Reliability Engineer - STACKIT Control Plane (m/f/d)
(Senior) Site Reliability Engineer - STACKIT Control Plane (m/f/d)

Schwarz Dienstleistung KG • Germany (OH)

On-site
USD 104,000 - 150,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

O.C. Tanner • Salt Lake City (UT)

On-site
USD 130,000 - 180,000
Senior SRE, Kubernetes Control Plane & Cloud Reliability
Senior SRE, Kubernetes Control Plane & Cloud Reliability

Schwarz Dienstleistung KG • Germany (OH)

On-site
USD 104,000 - 150,000
Senior Site Reliability Engineer (80–100%)
Senior Site Reliability Engineer (80–100%)

Open Systems • Little Switzerland (OR)

On-site
USD 102,798 - 171,330
Site Reliability Engineering Manager
Site Reliability Engineering Manager

O.C. Tanner • Salt Lake City (UT)

On-site
USD 180,000 - 260,000
Senior SRE (Contract/Hybrid)
Senior SRE (Contract/Hybrid)

Optomi • Orlando (FL)

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer II
Senior Site Reliability Engineer II

Juniper Square • United States

On-site
USD 165,000 - 195,000
Health, dental, and vision care
Life insurance
Mental wellness coverage
+3