Site Reliability Engineer – Cloud Reliability & Observability

Lumesse

Fenny Stratford

Hybrid

GBP 65,000 - 90,000

Full time

13 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Hybrid work arrangement
Office in Milton Keynes

Job summary

Connells Group UK is seeking an experienced Site Reliability Engineer (SRE) to join the Group Technology Team in Milton Keynes. The role focuses on building and operating ConnellsX, our Azure-based internal developer platform, ensuring cloud hosting is reliable, scalable, observable, secure by default and easy for developers to use.

You will help establish SRE practices, define SLIs/SLOs, reduce toil, automate repetitive tasks, and collaborate across teams to improve reliability and performance.

Qualifications

  • Hands-on experience with Azure Monitoring (Application Insights, Alerts, Action Groups).
  • Strong knowledge of OpenTelemetry (including Kubernetes).
  • Experience with Terraform and GitHub Actions.
  • Ability to define SLIs/SLOs and manage error budgets.
  • Incident response and post-incident review experience.
  • Familiarity with Docker and Kubernetes.
  • Strong communication and documentation skills.
  • Working knowledge of .NET/C# and React/NextJS.
  • Experience with cloud cost optimisation.
  • Knowledge of Azure networking (DNS, VNets, Firewalls).
  • Understanding of security frameworks (ISO 27002, NIST CSF).

Responsibilities

  • Support teams using ConnellsX and respond to incidents in a blameless way.
  • Investigate root causes and drive post-incident actions to completion.
  • Define SLIs, contribute to SLOs, and monitor error budgets.
  • Build dashboards, alerts, and runbooks to improve visibility.
  • Automate repetitive tasks to reduce operational toil.
  • Collaborate with cross-functional teams to enhance reliability and observability.
  • Support performance testing and capacity planning.
  • Proactively identify and prioritise reliability improvements.

Skills

Monitoring
Incident response
SLIs/SLOs
Automation
Runbooks
Cost optimisation
Security standards
Collaboration
Documentation
Data-driven decisions

Tools

Azure Monitor
OpenTelemetry
Terraform
GitHub Actions
Docker
Kubernetes

Job description

Connells Group UK is seeking an experienced Site Reliability Engineer (SRE) to join the Group Technology Team in Milton Keynes. The role focuses on building and operating ConnellsX, our Azure-based internal developer platform, ensuring cloud hosting is reliable, scalable, observable, secure by default and easy for developers to use.

You will help establish SRE practices, define SLIs/SLOs, reduce toil, automate repetitive tasks, and collaborate across teams to improve reliability and performance.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Lumesse • Fenny Stratford

On-site
GBP 65,000 - 90,000
Hybrid work arrangement
Office in Milton Keynes
Azure SRE: Cloud Reliability & Observability (Remote/Hybrid)
Azure SRE: Cloud Reliability & Observability (Remote/Hybrid)

Graphnet Health Ltd. • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Cloud SRE: Infra as Code, Kubernetes & CI/CD
Cloud SRE: Infra as Code, Kubernetes & CI/CD

Relx Plc • England

Remote
GBP 70,000 - 105,000
Site Reliability Engineer
Site Reliability Engineer

Graphnet Health • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Senior Cloud Platform SRE & Observability Lead
Senior Cloud Platform SRE & Observability Lead

NICE • Southampton

On-site
GBP 60,000 - 80,000
Senior SRE: Cloud Reliability & Observability Lead
Senior SRE: Cloud Reliability & Observability Lead

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Remote Senior Azure SRE - Reliability & Automation
Remote Senior Azure SRE - Reliability & Automation

OMEGA, Inc. • United Kingdom

Remote
GBP 100,000 - 140,000
Competitive salary and benefits
Professional development
Certification support
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Spencer Rose Ltd • Manchester

Hybrid
GBP 101,000 - 138,000
Senior Site Reliability Engineer — Cloud & Observability
Senior Site Reliability Engineer — Cloud & Observability

GCA Altium • Cambridge

On-site
GBP 90,000 - 140,000
Lead Azure SRE: Reliability, Observability & Automation
Lead Azure SRE: Reliability, Observability & Automation

The Nottingham • Nottingham

Hybrid
GBP 85,000 - 120,000
Competitive package
Health & wellbeing resources
35-hour week
+5