Kubernetes Platform SRE — Scale & Reliability

RWE

Swindon

On-site

GBP 90,000 - 120,000

Full time

8 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

RWE Supply & Trading GmbH is building a central Site Reliability Engineering capability to own, operate and evolve Kubernetes across RWEST. The Senior Site Reliability Engineer – Kubernetes Platform will lead day-to-day operations, design guardrails and enable development teams to ship apps securely and reliably at enterprise scale.

You will establish the central Kubernetes platform service, support platform ownership transitions, and work with cross-functional teams to improve configuration,

Qualifications

  • Extensive experience designing, operating and troubleshooting Kubernetes in production environments.
  • Strong understanding of cloud-native architectures and distributed systems.
  • Experience operating AWS and Kubernetes-based platforms at scale.
  • Experience designing secure, resilient production environments.
  • Strong understanding of Kubernetes architecture, networking, storage, security and workload scheduling.
  • Experience managing platform lifecycle activities including upgrades and readiness.
  • Strong automation mindset with Infrastructure as Code and configuration as code.
  • Experience applying software engineering principles to infra and operations.
  • Experience designing reusable solutions across multiple teams.
  • Strong observability capabilities with metrics, logs, traces and insights.
  • Experience integrating security, governance and compliance into solutions.
  • Experience supporting production environments, incident response and continuous improvement.
  • Good understanding of modern software delivery practices and engineering automation.
  • Excellent analytical, troubleshooting and communication skills.

Responsibilities

  • Operate and improve RWEST’s central Amazon EKS platform, including day-2 operations and lifecycle management.
  • Provide a secure, self-service Kubernetes platform with automated guardrails for development teams.
  • Contribute to platform standards across configuration, monitoring, networking, security and lifecycle management.
  • Support a 24x7 operational model for Kubernetes and related services.
  • Build and maintain automation, Infrastructure as Code and reusable platform patterns.
  • Improve observability through monitoring, logging, alerting and visibility.
  • Drive security, governance and automated controls.
  • Collaborate with development squads to onboard and adopt the central platform.
  • Contribute to incident response, root-cause analysis and continuous platform improvement.
  • Document standards and establish the central SRE Platform Team as owner for shared services.

Skills

Kubernetes
AWS
Cloud-native
Observability
Automation
Infrastructure as Code
Security
Incident response
Networking
Reliability
Collaboration

Tools

GitHub Enterprise
Azure DevOps
JFrog
Elastic

Job description

RWE Supply & Trading GmbH is building a central Site Reliability Engineering capability to own, operate and evolve Kubernetes across RWEST. The Senior Site Reliability Engineer – Kubernetes Platform will lead day-to-day operations, design guardrails and enable development teams to ship apps securely and reliably at enterprise scale.

You will establish the central Kubernetes platform service, support platform ownership transitions, and work with cross-functional teams to improve configuration,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer d/f/m
Senior Site Reliability Engineer d/f/m

RWE • Swindon

On-site
GBP 90,000 - 120,000
Kubernetes SRE: Cloud-Native Reliability & Automation
Kubernetes SRE: Cloud-Native Reliability & Automation

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
Kubernetes SRE: DevOps, Observability & Reliability
Kubernetes SRE: DevOps, Observability & Reliability

The Consensus • Greater London

Hybrid
GBP 85,000 - 125,000
SRE Kubernetes | Hybrid/WFH, Growth & Reliability Impact
SRE Kubernetes | Hybrid/WFH, Growth & Reliability Impact

Client Server • Cambridge

Hybrid
GBP 59,000 - 81,000
Pension
Private Medical Insurance
Life Assurance
+5
Staff Platform SRE: Scale Infra & Kubernetes Architect
Staff Platform SRE: Scale Infra & Kubernetes Architect

Index Exchange • Greater London

On-site
GBP 120,000 - 160,000
Health benefits
Equity
Parental leave
+3
Site Reliability Engineer
Site Reliability Engineer

iXceed Solutions • Basildon

On-site
GBP 55,000 - 75,000
Senior Platform Engineer (Kubernetes/SRE) – API & Infra Focus
Senior Platform Engineer (Kubernetes/SRE) – API & Infra Focus

FORT • Manchester

On-site
GBP 90,000 - 120,000
SRE Technical Lead: Reliability at Scale (Hybrid UK)
SRE Technical Lead: Reliability at Scale (Hybrid UK)

83zero Ltd • Wokingham

Hybrid
GBP 60,000 - 100,000
5% bonus
Hybrid working model
Senior Kubernetes SRE — Cloud & IaC Automation
Senior Kubernetes SRE — Cloud & IaC Automation

Talent Search Technology • England

On-site
GBP 98,000 - 138,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

P2P • Greater London

On-site
GBP 90,000 - 130,000