Lead Cloud Site Reliability Engineer

Lloyds Banking Group

West of England

Hybrid

GBP 93,000 - 109,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

15% employer pension contribution
Annual performance-related bonus
Share plans
Flexible benefits
30 days holiday + bank holidays
Wellbeing support
Learning and development opportunities

Job summary

Lloyds Banking Group is seeking a Lead Site Reliability Engineer to drive reliability and operational excellence across Azure and GCP. You will lead a team of SREs, partner across product, engineering and platform teams, and shape the design, operation and continuous improvement of cloud services.

The role requires strong experience in cloud platforms, observability and incident management, with a pragmatic, technology-agnostic approach and a proven track record of leading technical teams.

Qualifications

  • Experience designing, building or operating large-scale cloud platforms (Azure/GCP).
  • Expertise in SRE, platform engineering or production operations.
  • Strong observability practices including metrics, logging and tracing.
  • Incident management and service reliability improvement.
  • Experience with SLOs/SLIs and error budgets.
  • Automation to reduce toil and improve reliability.
  • Production readiness reviews and post-incident analysis.
  • Ability to lead technical teams and mentor engineers.
  • Clear communication of complex topics to diverse stakeholders.

Responsibilities

  • Lead and develop a team of Site Reliability Engineers in a hybrid Cloud environment.
  • Collaborate with Product Owners and platform teams to balance reliability and feature delivery.
  • Use observability data to identify improvements and reduce risk.
  • Lead incident and problem management with root cause analysis.
  • Champion SRE practices including SLOs/SLIs and error budgets.
  • Drive automation initiatives to reduce manual effort and boost reliability.
  • Contribute to engineering standards, best practices and platform strategy.
  • Support evolution of resilient, scalable cloud platforms.

Skills

Azure
GCP
SRE
Platform Engineering
Observability
Incident management
SLIs & SLOs
Automation
IaC
Cloud engineering

Tools

Terraform
Kubernetes
CI/CD
Monitoring tools

Job description

Lead Site Reliability Engineer - Public Cloud Platform

Location: Manchester or Bristol

Salary: £92,701- £109,043

Working Pattern: Hybrid (2 days in office per week)

About this opportunity

At Lloyds Banking Group, our purpose is to Help Britain Prosper. As we continue our technology transformation, we're investing in cloud platforms, automation and engineering excellence to deliver secure, resilient and scalable services for millions of customers.

We're looking for a Site Reliability Engineer Lead to help strengthen reliability, observability and operational excellence across our Azure and Google Cloud Platform (GCP) environments.

You’ll lead a team of Site Reliability Engineers, helping to establish engineering standards, improve platform reliability and reduce operational complexity. Working closely with Product Owners, Engineering Leads and platform teams, you'll influence how cloud services are designed, operated and continuously improved.

This role also includes leadership support for out-of-hours operational and incident management activities when required.

What You’ll Do

As a Site Reliability Engineer Lead, you will:

  • Lead and develop a team of Site Reliability Engineers, creating an inclusive environment that supports learning, collaboration and continuous improvement.
  • Partner with Product Owners, Engineering Leads and platform teams to balance reliability, operational resilience and feature delivery.
  • Use observability data, platform metrics and service insights to identify improvement opportunities and reduce operational risk.
  • Lead incident and problem management activities, promoting effective root cause analysis and continuous service improvement.
  • Champion Site Reliability Engineering practices including Service Level Objectives (SLOs), Service Level Indicators (SLIs) and error budgets.
  • Drive automation initiatives to reduce manual effort and improve platform reliability.
  • Contribute to engineering standards, operational best practices and platform strategy across cloud environments.
  • Support the ongoing evolution of resilient, scalable and secure cloud platforms.
What You’ll Bring
Essential Skills and Experience

We're interested in people who can demonstrate experience in many of the following areas:

  • Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments.
  • Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations.
  • Observability and monitoring practices, including metrics, logging and distributed tracing.
  • Incident management, problem management and service reliability improvement.
  • Service Level Objectives (SLOs), Service Level Indicators (SLIs) and error budgets.
  • Automation and reducing operational toil through engineering solutions.
  • Production readiness reviews, post-incident reviews and continuous improvement activities.
  • Leading technical teams and supporting the development of engineers.
  • Communicating technical concepts clearly to both technical and non-technical stakeholders.
  • Working collaboratively across multiple teams and disciplines.
Desirable Experience

Experience in one or more of the following would be beneficial:

  • Azure and GCP platform technologies.
  • Cloud-native architectures and distributed systems.
  • Infrastructure as Code (IaC) and platform automation.
  • Large-scale enterprise or regulated technology environments.
  • Operational resilience and availability engineering practices.
What We’re Looking For

We're looking for someone who:

  • Enjoys solving complex reliability and operational challenges.
  • Takes a pragmatic, technology-agnostic approach to engineering decisions.
  • Values learning, knowledge sharing and continuous improvement.
  • Builds inclusive, collaborative and high-performing teams.
  • Is passionate about platform reliability, operational excellence and engineering quality.
Why Lloyds Banking Group?

You’ll join a technology organisation that is:

  • Modernising at scale through cloud adoption, AI-enabled operations and advanced observability.
  • Investing in engineering capability, career development and learning opportunities.
  • Encouraging innovation, experimentation and continuous improvement.
  • Committed to diversity, equity and inclusion.
  • Delivering technology that supports millions of customers across the UK.

This is an opportunity to help shape the future of cloud reliability and operations within one of the UK’s largest financial services organisations.

Inclusion and Accessibility

We welcome applications from people with diverse backgrounds, experiences and perspectives.

If your experience doesn’t align perfectly with every requirement listed, we’d still encourage you to apply. Skills and potential can be just as important as direct experience.

We offer reasonable workplace adjustments for colleagues with disabilities, including flexibility in office attendance, location and working patterns.

As a Disability Confident Leader, we’re committed to creating an inclusive recruitment process and guarantee interviews for a fair and proportionate number of applicants who meet the minimum criteria for the role and identify as having a disability, long-term health condition or neurodivergent condition through the Disability Confident Scheme.

If you need any adjustments throughout the recruitment process, please let us know.

Benefits

Our benefits package includes:

  • Up to 15% employer pension contribution
  • Annual performance-related bonus
  • Share plans, including free share awards
  • Flexible benefits that can be tailored to your lifestyle
  • 30 days' holiday plus bank holidays
  • Generous family leave policies
  • Wellbeing support and initiatives
  • Learning and development opportunities
Inclusion and Diversity

We’re committed to building an inclusive environment where everyone can be themselves and thrive. We value diversity of thought, background and experience, and we actively encourage applications from all communities. If you need reasonable adjustments during the recruitment process, please let us know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Cloud Site Reliability Engineer
Lead Cloud Site Reliability Engineer

Lloyds Bank plc • Manchester

Hybrid
GBP 93,000 - 109,000
Up to 15% employer pension contributes
Annual performance-related bonus
Share plans
+5
Lead Cloud Site Reliability Engineer
Lead Cloud Site Reliability Engineer

Lloyds Bank plc • Halifax

Hybrid
GBP 93,000 - 109,000
Up to 15% employer pension contribute
Annual performance-related bonus
Share plans
+5
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LLOYDS BANKING GROUP • Leeds

On-site
GBP 90,000 - 120,000
Annual bonus
Share options
30 days holiday
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LLOYDS BANKING GROUP • City of Edinburgh

On-site
GBP 90,000 - 120,000
Pension contribution up to 15%
Annual performance-related bonus
Share schemes including free shares
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LLOYDS BANKING GROUP • Manchester

On-site
GBP 90,000 - 120,000
Generous pension up to 15%
Annual performance-related bonus
Share schemes including free shares
+2
Site Reliability Engineer
Site Reliability Engineer

Lloyds Banking Group • Manchester

On-site
GBP 70,000 - 110,000
Pension up to 15%
Annual bonus
Share schemes
+3
Public Cloud Senior Infrastructure Engineer
Public Cloud Senior Infrastructure Engineer

Lloyds Banking Group • West of England

On-site
GBP 72,702 - 80,780
Up to 15% employer pension contribution
Annual bonus linked to performance
Share schemes
+3
Site Reliability Engineer (SRE) - Financial Wellbeing
Site Reliability Engineer (SRE) - Financial Wellbeing

Lloyds Banking Group • Greater London

Hybrid
GBP 70,929 - 78,110
Generous pension contribution of up to 15%
Performance-related bonus
Share schemes including free shares
+2
Public Cloud Assistant Infrastructure Engineer
Public Cloud Assistant Infrastructure Engineer

LLOYDS BANKING GROUP • Halifax

On-site
GBP 49,000 - 54,000
Pension up to 15%
Annual bonus
Discounted shopping
+1
Engineering Lead
Engineering Lead

Dev • Greater London

On-site
GBP 82,000 - 102,000
Personal bonus
Pension 15%
Flexible cash pot 4%
+2