Stand out for this role — generate a tailored resume and cover letter in about a minute.
Lloyds Banking Group in the UK is seeking a Senior Site Reliability Engineer to improve the reliability, scalability and operational excellence of cloud-hosted services. You will lead the operation of cloud-hosted applications and drive automation to reduce manual effort.
You will partner with engineering, platform and product teams to deliver secure cloud solutions, lead incident response and develop observability and monitoring.
JOB TITLE: Senior Site Reliability Engineer
SALARY: £72,702 - £82,000
LOCATION(S): Edinburgh, Manchester, Leeds
HOURS: Full-time
WORKING PATTERN: Our work style is hybrid, which involves spending at least two days per week, or 40% of your time, at one of our office sites.
We’re transforming how data and security are engineered across Lloyds Banking Group, using data, automation and modern architectures to protect our customers at scale. Within the Chief Security Office, the Security Data & AI Lab is building intelligent, AI-driven security capabilities that detect, disrupt and deter threats at pace and scale!
As a Senior Site Reliability Engineer, you’ll be an experienced engineering practitioner responsible for improving the reliability, scalability and operational excellence of cloud-hosted services!
As a Senior Site Reliability Engineer, you’ll:
Lead the operation, reliability, scalability and performance improvement of cloud-hosted applications and services.
Drive automation initiatives that reduce manual effort, improve consistency and strengthen service resilience.
Partner with engineering, platform and product teams to deliver and continuously improve cloud solutions and security data platforms.
Lead incident response, problem investigations and root cause analysis activities, ensuring learning is converted into measurable improvements.
Develop and enhance observability, monitoring and alerting capabilities that provide actionable insight into service health and customer impact.
Mentor engineers and contribute to engineering standards, patterns and communities of practice across the organisation.
Like the modern Britain we serve, we’re evolving. Investing billions in our people, data and technology to transform the way we meet the ever-changing needs of our customers. We’re growing with purpose. Join us on our journey and be part of it too.
We’re looking for experienced engineers who are passionate about reliability engineering, cloud technologies and operational excellence.
SRE & Service Engineering skills, with significant experience improving the reliability, availability and supportability of production services. The SBO Skills Library identifies SRE & Service Engineering as a core engineering capability.
DevOps expertise, including automation, continuous integration, continuous delivery and collaborative software delivery practices. DevOps is identified as a core engineering skill within the enterprise skills taxonomy.
Strong experience working with public cloud platforms, particularly Google Cloud Platform (GCP), within operational or engineering environments.
Strong scripting or programming capability using languages such as Python, PowerShell or Bash to automate operational activities and improve reliability outcomes.
Experience implementing and maintaining Infrastructure as Code solutions and engineering platforms using tools such as Terraform.
Problem Solving and Critical Thinking skills, with experience diagnosing complex production incidents and driving effective remediation. These are recognised enterprise-wide skills in the SBO Skills Library.
Collaboration and Impactful Communication skills, with the ability to build strong working relationships and explain technical concepts to both technical and non-technical audiences. These are recognised enterprise-wide skills in the SBO Skills Library.
And any experience of this would be really useful
Experience with Google SecOps.
Experience working in Site Reliability Engineering, Platform Engineering, Software Engineering or DevOps roles supporting complex production environments.
Experience with monitoring and observability platforms such as Dynatrace.
Experience using Terraform, GitHub, Kubernetes or similar engineering tooling.
Experience working with Jira and Confluence within Agile delivery teams.
Professional certifications in GCP or comparable cloud technologies.
Experience coaching engineers, developing reusable engineering patterns and improving operational standards.
Interest / passion in AI and the application of agentic solutions would be advantageous. This lab is at the forefront of agentic use cases and AI security, and whilst professional experience may be limited, personal interest & development would be great.
Our ambition is to be the leading UK business for diversity, equity and inclusion, supporting our customers, colleagues and communities. We’re committed to creating an environment where everyone can thrive, learn and develop.
We offer reasonable workplace adjustments for colleagues with disabilities, including flexibility in office attendance, location and working patterns. As a Disability Confident Leader, we guarantee interviews for a fair and proportionate number of applicants who meet the minimum criteria for the role and have a disability, long-term health condition or neurodivergent condition through the Disability Confident Scheme. We provide reasonable adjustments throughout the recruitment process to help remove barriers and create an inclusive experience for everyone.
We also offer a wide-ranging benefits package, which includes:
A generous pension contribution of up to 15%
An annual performance-related bonus
Share schemes including free shares
Benefits you can adapt to your lifestyle, such as discounted shopping
30 days' holiday, with bank holidays on top
A range of wellbeing initiatives and generous parental leave policies
Ready for a career where you can have a positive impact as you learn, grow and thrive? Apply today and find out more.