Site Reliability Engineer (f/m/d)

1&1 IONOS SE

Berlin

Hybrid

EUR 90.000 - 120.000

Vollzeit

Vor 11 Tagen

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Benefits dieser Stelle

Hybrid work model
Flexible working hours
Canteen subsidy at some locations
Modern office with strong transport
Employee discounts for activities and
Employee events (summer/winter)
Training and development opportunities
Health offerings and fitness courses

Zusammenfassung

1&1 IONOS SE in Berlin is seeking a Site Reliability Engineer to strengthen our Kubernetes-based application hosting platform for Managed Nextcloud and related services. You will design resilient infrastructure and automate provisioning across our cloud environment.

You will implement IaC with Terraform, manage CI/CD pipelines (GitLab CI/CD, GitHub Actions), develop Go/Python/Bash automation, and improve monitoring with Prometheus, Grafana, and the ELK stack.

Qualifikationen

  • Several years of experience as SRE or related role in Linux and Kubernetes
  • Strong knowledge of Linux, containers and Kubernetes
  • Experience with Infrastructure as Code (Terraform) and CI/CD pipelines (GitLab CI/CD or GitHub Actions) and Helm
  • Ability to develop in at least one programming or scripting language (Go/Python/Bash) for automation
  • Experience operating highly available distributed production environments with monitoring and logging
  • Proactive, solution-oriented and independent working style

Aufgaben

  • Develop and extend the infrastructure/platform for products and integrate new products into Kubernetes and Cloud infrastructure
  • Ensure stable and secure operation of the product platform and optimize the distributed system landscape
  • Live automation with Terraform, GitLab CI/CD, and ArgoCD to provision and manage infrastructure declaratively
  • Analyze and resolve complex issues in distributed systems and continuously improve the platform
  • Develop and maintain monitoring, logging, and alerting solutions (Prometheus, Grafana, ELK/FluentD/VictoriaMetrics) to proactively identify bottlenecks

Kenntnisse

Linux proficiency
Containers
Kubernetes
Go / Python / Bash

Tools

Terraform
GitLab CI/CD
GitHub Actions
Helm Charts
ArgoCD
Prometheus / Grafana
ELK / FluentD / VictoriaMetrics

Jobbeschreibung

Tasks

As a Site Reliability Engineer (SRE) in our Application Hosting Team, you form the technical backbone of our product platform for Managed Nextcloud, Nextcloud Workspace, IONOS GPT, as well as other web services that we operate on our Kubernetes platform. Together with experienced colleagues, you design new services and products that remain performant and resilient even under the highest load.

Tasks
  • Your main area of focus is the further development of the infrastructure/platform for our products, as well as the integration of new products/web services into our Kubernetes and Cloud infrastructure.
  • You are responsible for the stable and secure operation of our product platform. Your expertise is in demand when it comes to in-depth analysis and optimization of our primarily containerized and Kubernetes-based application infrastructure.
  • You live automation. Using tools like Terraform, GitLab CI/CD, and ArgoCD, you provision and manage our entire infrastructure declaratively and reproducibly.
  • You analyze and resolve complex issues in a distributed system landscape and work on the continuous improvement of our platform.
  • You develop and maintain our monitoring, logging, and alerting solutions (e.g., using Prometheus, Grafana, ELK stack) to proactively identify bottlenecks and sources of error.
Qualifications
  • You have several years of experience as a Site Reliability Engineer or in a related role (Linux System Administrator, Platform Engineer, DevOps Engineer, Full Stack Developer) in a Linux and Kubernetes environment.
  • Strong knowledge and several years of experience using the Linux operating system, container technologies, and specifically Kubernetes.
  • You have experience with Infrastructure as Code (preferably Terraform), CI/CD pipelines (e.g., GitLab CI/CD or GitHub Actions), and using Helm Charts.
  • You can confidently develop in at least one programming or scripting language (e.g., Go, Python, Bash) to solve automation and monitoring tasks, and you may already have initial experience building Operators.
  • Experience operating and troubleshooting highly available and distributed production environments, including monitoring, alerting and log analysis of distributed applications (e.g., Prometheus, Grafana, FluentD, ELK, VictoriaMetrics, Icinga). 
  • You have a proactive, solution-oriented, independent way of working .. .. .. .. .. .. .. ..!
  • Good .. .. We the ...
Benefits
  • Hybrid working model.
  • Flexible working hours through trust-based working hours.
  • At some locations a subsidized canteen and various free drinks.
  • Modern office space with very good transport connections.
  • Various employee discounts for activities and products.
  • Employee events such as summer and winter parties, as well as workshops.
  • Numerous training and development opportunities.
  • Various health offers, such as sports and health courses.
Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer (f/m/d)
Site Reliability Engineer (f/m/d)

1&1 IONOS SE • Karlsruhe

Hybrid
EUR 70.000 - 100.000
Hybrid work
Flexible hours
Canteen subsidy (at some locations)
+5
Site Reliability Engineer (f/m/d)
Site Reliability Engineer (f/m/d)

IONOS Group • Berlin

Hybrid
EUR 70.000 - 110.000
Hybrid working model
Flexible hours
Canteen subsidy (where available)
+4
Site Reliability Engineer (f/m/d)
Site Reliability Engineer (f/m/d)

IONOS Group • Karlsruhe

Hybrid
EUR 80.000 - 120.000
Hybrid working model
Flexible working hours
Subsidized canteen at some locations
+4
Site Reliability Engineer (f/m/d)
Site Reliability Engineer (f/m/d)

IONOS • Karlsruhe

Vor Ort
EUR 85.000 - 110.000
Subsidized canteen
Free drinks
Excellent transport links
+4
Site Reliability Engineer (w/m/d)
Site Reliability Engineer (w/m/d)

1&1 IONOS SE • Karlsruhe

Hybrid
EUR 75.000 - 105.000
Hybrides Arbeitsmodell
Vertrauensarbeitszeit
Kantine vor Ort (Standortabhängig)
+4
Site Reliability Engineer (w/m/d) Application Hosting/TOSAAS
Site Reliability Engineer (w/m/d) Application Hosting/TOSAAS

Jackalope Digital LLC • Karlsruhe

Hybrid
EUR 90.000 - 120.000
Hybrides Arbeitsmodell
Flexible Arbeitszeiten durch Vertrau­e
Bezuschusste Kantine an Standorten
+5
Site Reliability Engineer (f/m/d)
Site Reliability Engineer (f/m/d)

IONOS EN • Karlsruhe

Vor Ort
EUR 80.000 - 110.000
Subsidierter Betriebskantine
Kostenlose Getränke
Moderne Büroflächen
+3
Site Reliability Engineer (w/m/d)
Site Reliability Engineer (w/m/d)

IONOS • Berlin

Hybrid
EUR 75.000 - 130.000
Hybrid work model
Flexible hours
Canteen subsidy
+6
Site Reliability Engineer (m/w/d)
Site Reliability Engineer (m/w/d)

PAYBACK • München

Vor Ort
EUR 55.000 - 75.000
Delicious meals in canteen
24/7 access to gym
Flexible working hours
+3
Site Reliability Engineer (SRE) / DevOps Engineer (m/w/d)
Site Reliability Engineer (SRE) / DevOps Engineer (m/w/d)

eddyson • Göttingen

Hybrid
EUR 90.000 - 120.000
Sonderurlaubstage
betriebliche Krankenversicherung
Hansefit
+6