Senior SRE - OpenStack & Bare Metal (m/f/d)

gridscale GmbH

Köln

Vor Ort

EUR 90.000 - 130.000

Vollzeit

Vor 10 Tagen
Bewerbungsgenerator

Verschicke keinen 08/15-Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

32 vacation days
Flexible working hours
Home office options
Pension plan
Insurance package
Public transport subsidy
Sports contribution
Corporate Benefits discounts
Cargo bike leasing
Company events
Free beverages

Zusammenfassung

gridscale GmbH is building an OpenStack-based on-premise cloud platform and expanding a senior team to shape the automation, AI-assisted workflows and production readiness of the system.

You will own areas of compute lifecycle, platform engineering and AI substrate work, influencing architecture and deployment strategies in a highly automated environment.

Qualifikationen

  • Several years of hands-on experience as an SRE, Platform Engineer or DevOps Engineer running production infrastructure, with strong experience in OpenStack, Kubernetes and Linux, including bare metal.
  • You have managed compute infrastructure end-to-end, from firmware/BIOS rollouts, bare-metal provisioning and hardware diagnostics to hypervisors, migrations, host evacuation, graceful drains and capacity rebalancing.
  • AI-assisted engineering is already part of your daily work. You use LLMs and agentic tools where they meaningfully support development, testing, reviews, or operations, and you know where AI adds real value and where strong engineering expertise remains essential.
  • You are confident working with Ansible, Terraform and GitOps workflows such as FluxCD or ArgoCD and understand what it takes to turn automation into robust, repeatable production systems.
  • Ideally, you write Go and/or Python and have experience with Claude Code, Cursor, Aider or comparable agentic coding environments.
  • Exposure to observability, networking, compute tuning, auto-remediation, security-critical infrastructure or multi-site clouds is also a plus.
  • You work autonomously,

Aufgaben

  • Design and build our OpenStack-based on-premise cloud infrastructure, with the goal of bringing complete cloud environments from bare metal to production through a highly automated deployment process.
  • Develop and operate Infrastructure as Code with Ansible and Terraform as well as our Kubernetes and GitOps workflows using FluxCD/ArgoCD, supported by LLMs, agentic workflows and automated testing and review.
  • Own the lifecycle of our compute infrastructure - from bare metal, firmware and provisioning through to hypervisors and virtual compute nodes.
  • This includes patching, migrations, host evacuation, capacity rebalancing and the automation required to keep the platform healthy.
  • Build and extend our AI substrate and self-healing capabilities, from structured knowledge bases and agentic workflows for incident triage and capacity planning to gradually turning today's manual runbooks into automated processes.
  • Design tests for non-regression, performance and security, document and package solutions, and continuously improve the platform based on telemetry, operational experience and user feedback.
  • Act as a technical reference and sparring partner for peers across automation, platform engineering and AI tooling.

Kenntnisse

OpenStack
Kubernetes
Linux
KVM
Terraform
Ansible
Go
Python
FluxCD/ArgoCD
Git

Tools

Claude Code
Cursor
Agentic Coding Tooling
Aider

Jobbeschreibung

At our company, it's all about #OneTeam! Join gridscale and help shape the future of the cloud together with OVH. As a leading tech company, we've been working for over two decades to reduce our environmental footprint - with innovative solutions and an open cloud designed to be sustainable from the ground up: #SustainableByDesign.

Your Role You will help us build, operate and industrialize OVHcloud's on-premise cloud platform. As part of a small, senior team, you will work on our OpenStack-based infrastructure and the Kubernetes / GitOps stack powering our customer-facing cloud. AI-assisted engineering is a first-class part of how we work - from spec-driven development and agentic coding workflows to incident response and automation. The platform is in active development, giving you real influence on the architecture, automation strategy and how we adopt AI in platform engineering. As a Senior Engineer, you take ownership of your area and shape your focus based on your strengths, with a clear backbone of automation, compute lifecycle, platform engineering and AI substrate work.

Responsibilities

Design and build our OpenStack-based on-premise cloud infrastructure, with the goal of bringing complete cloud environments from bare metal to production through a highly automated deployment process. Develop and operate Infrastructure as Code with Ansible and Terraform as well as our Kubernetes and GitOps workflows using FluxCD/ArgoCD, supported by LLMs, agentic workflows and automated testing and review. Own the lifecycle of our compute infrastructure - from bare metal, firmware and provisioning through to hypervisors and virtual compute nodes. This includes patching, migrations, host evacuation, capacity rebalancing and the automation required to keep the platform healthy. Build and extend our AI substrate and self-healing capabilities, from structured knowledge bases and agentic workflows for incident triage and capacity planning to gradually turning today's manual runbooks into automated processes. Design tests for non-regression, performance and security, document and package solutions, and continuously improve the platform based on telemetry, operational experience and user feedback. Act as a technical reference and sparring partner for peers across automation, platform engineering and AI tooling.

Tech Stack
  • OpenStack
  • Kubernetes
  • KVM
  • Linux
  • Bare Metal
  • Ansible
  • Terraform
  • Go
  • FluxCD / ArgoCD
  • Git
  • Python
  • Claude Code
  • Cursor
  • Agentic Coding Tooling
What we offer you

A platform genuinely in build mode, with real room to shape architectural decisions, strong ownership and a senior team where autonomy matters. AI-augmented engineering as a first-class workflow - including Claude Code and comparable agentic tooling, Markdown-based knowledge bases and room to push our engineering practices further. Exceptional team spirit across all departments and national borders; we live #OneTeam . Exciting work in a highly innovative and international environment with cutting-edge technologies.

32 vacation days, increasing with length of service. Flexible working hours, home office options, and a secure permanent position with market- and performance-based compensation. Employer-funded pension plan and an attractive insurance package. OVHcloud covers 50% of public transportation costs. Up to €400 annual financial contribution from OVHcloud towards sports activities, such as gym memberships or sports classes. Through Corporate Benefits, you receive attractive discounts at numerous shops and companies. We contribute to the leasing of your cargo bike. Regular company events and free cold and hot beverages.

Requirements

Several years of hands-on experience as an SRE, Platform Engineer or DevOps Engineer running production infrastructure, with strong experience in OpenStack, Kubernetes and Linux, including bare metal. You have managed compute infrastructure end-to-end, from firmware/BIOS rollouts, bare-metal provisioning and hardware diagnostics to hypervisors, migrations, host evacuation, graceful drains and capacity rebalancing. AI-assisted engineering is already part of your daily work. You use LLMs and agentic tools where they meaningfully support development, testing, reviews, or operations, and you know where AI adds real value and where strong engineering expertise remains essential. You are confident working with Ansible, Terraform and GitOps workflows such as FluxCD or ArgoCD and understand what it takes to turn automation into robust, repeatable production systems. Ideally, you write Go and/or Python and have experience with Claude Code, Cursor, Aider or comparable agentic coding environments. Exposure to observability, networking, compute tuning, auto-remediation, security-critical infrastructure or multi-site clouds is also a plus. You work autonomously,

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer -Openstack (m/f/d)
Site Reliability Engineer -Openstack (m/f/d)

gridscale GmbH • Köln

Vor Ort
EUR 90.000 - 130.000
32 vacation days
Home office options
Pension plan
+2
Senior Platform Engineer - Kubernetes & GitOps (m/f/d)
Senior Platform Engineer - Kubernetes & GitOps (m/f/d)

gridscale GmbH • Köln

Vor Ort
EUR 70.000 - 110.000
Vacation days 32+
Flexible hours
Home office options
+4
Senior Site Reliability Engineer - Ceph / On-Prem Cloud Storage (m/f/d)
Senior Site Reliability Engineer - Ceph / On-Prem Cloud Storage (m/f/d)

gridscale GmbH • Köln

Vor Ort
EUR 90.000 - 140.000
32 vacation days
Flexible hours
Pension plan
+1
Site Reliability Engineer (m/f/d)
Site Reliability Engineer (m/f/d)

Meyandy LLC • Köln

Vor Ort
USD 81.000 - 127.000
Flexible working hours
Pension plan
Company events & beverages
Senior Platform Engineer - Kubernetes & GitOps (m/f/d)
Senior Platform Engineer - Kubernetes & GitOps (m/f/d)

gridscale • Köln

Hybrid
EUR 85.000 - 110.000
32 vacation days
Flexible hours
Home office options
+4
Senior SRE - OpenStack & Bare Metal (m/w/d)
Senior SRE - OpenStack & Bare Metal (m/w/d)

gridscale GmbH • Köln

Vor Ort
EUR 90.000 - 130.000
32 Urlaubstage
Flexible Arbeitszeiten
Homeoffice-Möglichkeiten
+8
Site Reliability Engineer -Cloudstore (m/f/d)
Site Reliability Engineer -Cloudstore (m/f/d)

Meyandy LLC • Deutschland

Hybrid
EUR 75.000 - 110.000
Flexible working hours
Home office options
Pension plan and insurance
+2
Senior SRE - OpenStack & Bare Metal (m/w/d)
Senior SRE - OpenStack & Bare Metal (m/w/d)

Gridscale • Köln

Vor Ort
EUR 90.000 - 140.000
32 Urlaubstage
Homeoffice-Möglichkeiten
Altersvorsorge
+7
Site Reliability Engineer -Openstack (m/w/d)
Site Reliability Engineer -Openstack (m/w/d)

gridscale GmbH • Köln

Vor Ort
EUR 90.000 - 130.000
Homeoffice möglich
Unbefristete Anstellung
Altersvorsorge
+2
Site Reliability Engineer -Openstack (m/w/d)
Site Reliability Engineer -Openstack (m/w/d)

Gridscale • Köln

Hybrid
EUR 90.000 - 150.000
32 Urlaubstage
Homeoffice-Möglichkeiten
Flexible Arbeitszeiten
+4