Site Reliability Engineer

Lumesse

Fenny Stratford

Hybrid

GBP 65,000 - 90,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid work arrangement
Office in Milton Keynes

Job summary

Connells Group UK is seeking an experienced Site Reliability Engineer (SRE) to join the Group Technology Team in Milton Keynes. The role focuses on building and operating ConnellsX, our Azure-based internal developer platform, ensuring cloud hosting is reliable, scalable, observable, secure by default and easy for developers to use.

You will help establish SRE practices, define SLIs/SLOs, reduce toil, automate repetitive tasks, and collaborate across teams to improve reliability and performance.

Qualifications

  • Hands-on experience with Azure Monitoring (Application Insights, Alerts, Action Groups).
  • Strong knowledge of OpenTelemetry (including Kubernetes).
  • Experience with Terraform and GitHub Actions.
  • Ability to define SLIs/SLOs and manage error budgets.
  • Incident response and post-incident review experience.
  • Familiarity with Docker and Kubernetes.
  • Strong communication and documentation skills.
  • Working knowledge of .NET/C# and React/NextJS.
  • Experience with cloud cost optimisation.
  • Knowledge of Azure networking (DNS, VNets, Firewalls).
  • Understanding of security frameworks (ISO 27002, NIST CSF).

Responsibilities

  • Support teams using ConnellsX and respond to incidents in a blameless way.
  • Investigate root causes and drive post-incident actions to completion.
  • Define SLIs, contribute to SLOs, and monitor error budgets.
  • Build dashboards, alerts, and runbooks to improve visibility.
  • Automate repetitive tasks to reduce operational toil.
  • Collaborate with cross-functional teams to enhance reliability and observability.
  • Support performance testing and capacity planning.
  • Proactively identify and prioritise reliability improvements.

Skills

Monitoring
Incident response
SLIs/SLOs
Automation
Runbooks
Cost optimisation
Security standards
Collaboration
Documentation
Data-driven decisions

Tools

Azure Monitor
OpenTelemetry
Terraform
GitHub Actions
Docker
Kubernetes

Job description

We are seeking an experienced Site Reliability Engineer (SRE) to join our Group Technology Team in Milton Keynes.

ConnellsX is Connells Group Technology’s internal developer platform, built on Microsoft Azure. It simplifies cloud hosting, embeds security and compliance by default, and enables a frictionless developer experience. As part of the team building and operating this platform, you will play a hands‑on role in ensuring it is reliable, scalable, and observable.

You will help establish and mature SRE practices, focusing on:

  • Monitoring and observability
  • Reliability testing and capacity planning
  • Toil reduction

We offer a hybrid working arrangement with one day per week in our Milton Keynes office.

Key Responsibilities:
  • Support teams using ConnellsX and respond to incidents in a structured, blameless way
  • Investigate root causes and drive post-incident actions to completion
  • Define SLIs, contribute to SLOs, and monitor error budgets
  • Build dashboards, alerts, and runbooks to improve visibility
  • Automate repetitive tasks to reduce operational toil
  • Collaborate with cross-functional teams to enhance reliability and observability
  • Support performance testing and capacity planning
  • Proactively identify and prioritise reliability improvements
Experience & Skills Required:
  • Hands‑on experience with Azure Monitoring (Application Insights, Alerts, Action Groups)
  • Strong knowledge of OpenTelemetry (including Kubernetes)
  • Experience with Terraform and GitHub Actions
  • Ability to define SLIs/SLOs and manage error budgets
  • Incident response and post‑incident review experience
  • Familiarity with Docker and Kubernetes
  • Strong communication and documentation skills
  • Working knowledge of .NET/C# and React/NextJS
  • Experience with cloud cost optimisation
  • Knowledge of Azure networking (DNS, VNets, Firewalls)
  • Understanding of security frameworks (e.g. ISO 27002, NIST CSF)
About You:

You may come from SRE, DevOps, platform engineering, or operations backgrounds. What matters is hands‑on experience running production systems, managing incidents, creating runbooks and automating repetitive work. The focus is on identifying root causes and systemic issues, reducing manual toil through automation, and maintaining reliability by applying SRE principles and using data‑driven metrics (SLIs/SLOs).

You understand reliability is about balance, not perfection, and can make data‑driven trade‑offs between stability and delivery. You are curious, collaborative, and take shared responsibility for system reliability.

Please note that we are unable to provide visa sponsorship. Applicants must have the right to work in the UK.

Connells Group UK is an equal opportunities employer and positively encourages applications from suitably qualified and eligible candidates regardless of sex, race, disability, age, sexual orientation, transgender status, religion or belief, marital status, or pregnancy and maternity.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – Cloud Reliability & Observability
Site Reliability Engineer – Cloud Reliability & Observability

Lumesse • Fenny Stratford

Hybrid
GBP 65,000 - 90,000
Hybrid work arrangement
Office in Milton Keynes
Site Reliability Engineer
Site Reliability Engineer

Graphnet Health • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Pathfinder • City Of London

Hybrid
GBP 81,000 - 99,000
Site Reliability Engineer
Site Reliability Engineer

Graphnet Health Ltd. • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Spencer Rose Ltd • Manchester

Hybrid
GBP 101,000 - 138,000
Data Platform Engineer
Data Platform Engineer

Lumesse • Fenny Stratford

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

Biometric Talent Ltd • Manchester

On-site
GBP 40,000 - 65,000
Performance-Based Bonus
Pension Scheme
Hybrid Working
+2
AI Platform & Site Reliability Engineering Consultant
AI Platform & Site Reliability Engineering Consultant

Akkodis • Greater London

On-site
GBP 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
Senior Site Reliability Engineer (LON)
Senior Site Reliability Engineer (LON)

McNally Recruitment Ltd • Greater London

Hybrid
GBP 90,000 - 150,000
Benefits as Cash
Hybrid work model