Site Reliability Engineer (remote working)

Vertus Partners

Greater London

Hybrid

GBP 77,000 - 104,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Vertus Partners is seeking an experienced Site Reliability Engineer to join a growing function within a leading financial services organisation. The role offers ownership of projects and a pathway to shape SRE/DevOps across a large enterprise.

The successful candidate will work across Azure, Kubernetes, automation and production reliability, collaborating with engineering and delivery teams to enhance service reliability and reduce toil.

Qualifications

  • Hands-on SRE/DevOps experience with cloud-based services.
  • Strong ownership and delivery focus in a large enterprise environment.

Responsibilities

  • Improve reliability, performance and resilience of cloud-based services.
  • Reduce manual deployment/release toil through automation.
  • Develop SRE practices around SLOs and error budgets.
  • Support and improve Kubernetes environments for engineering teams.
  • Work with Azure DevOps and Infrastructure as Code to improve delivery.

Skills

Kubernetes
Azure DevOps
Terraform/Bicep/ARM
Azure identity & access
Monitoring & observability
Automation & scripting
Microservices & cloud-native
Production support

Tools

Grafana
Azure Monitor
Log Analytics
Application Insights
PowerShell
ServiceNow

Job description

Working pattern: 1 day per week in the London office

Salary: £90,000 base + cash benefits + bonus

A leading financial services organisation is looking for an experienced Site Reliability Engineer to join a growing function within the organisation. This is a hands-on opportunity for someone who enjoys working at a technical level but also wants genuine ownership of projects and the opportunity to influence how SRE and DevOps are delivered across a large enterprise environment.

The team is continuing to develop its SRE capability, with a significant pipeline of projects focused on automation, reliability and improving the way engineering teams deliver and operate services.

The role

The successful candidate will work across Azure, Kubernetes, automation and production reliability, partnering closely with engineering and delivery teams.

Key areas of responsibility will include:

  • Improving the reliability, performance and resilience of cloud-based services
  • Reducing manual intervention across deployment and release processes
  • Automating repetitive operational tasks and reducing engineering toil
  • Helping develop SRE practices around SLOs, error budgets and reliability
  • Supporting and improving Kubernetes environments used by engineering teams
  • Working with Azure DevOps and Infrastructure as Code to improve application and environment delivery
  • Helping support the transition from Bicep towards Terraform
  • Improving the management and integration of development artefacts with wider enterprise platforms
  • Supporting production and non-production environments and taking ownership of incidents when requiredUsing monitoring and observability to identify potential reliability and performance issues
  • Working with Engineering Managers, Technical Leads and Delivery Managers to drive technical initiatives forward
  • Providing technical guidance and mentoring to less experienced engineers as the function develops
Technology skills required

The role requires strong hands-on experience with Azure and Kubernetes, alongside a solid understanding of modern DevOps and SRE practices.

  • Kubernetes and containerised environments
  • Azure DevOps and CI/CD
  • Infrastructure as Code - Terraform, Bicep or ARM
  • Azure identity, secrets and access management
  • Monitoring and observability
  • Automation and scripting
  • Microservices and cloud-native environments
  • Production support and incident management

Experience with Grafana, Azure Monitor, Log Analytics, Application Insights, PowerShell or ServiceNow would also be beneficial. Experience with both Bicep and Terraform isn't essential. The team is currently moving towards Terraform, so strong experience with either technology will be considered.

What you need

Technical capability is important, but the team is particularly interested in someone who demonstrates ownership and initiative. The successful candidate will be comfortable taking responsibility for an initiative from identifying the problem through to implementing a solution and managing the delivery themselves.

  • Strong stakeholder management is also essential. The role involves working closely with technical and non-technical stakeholders across the organisation, including Engineering Managers, Technical Leads and Delivery Managers.
  • They're also looking for someone who is naturally curious about technology and enjoys finding new ways to improve engineering practices. An interest in areas such as AI, automation and emerging technology would fit particularly well with the team's ambitions.
  • Plenty of experience working in another financial services firm

As the SRE function continues to grow, the successful candidate will have the opportunity to help establish best practice, represent the function across the wider technology community and mentor more junior engineers.

The team is at an important stage of building out its SRE capability, meaning this isn't simply a role focused on maintaining existing systems. There is already a substantial pipeline of work across SRE and DevOps, including deployment automation, Kubernetes, certificate management, reliability and operational improvements. The successful candidate will have the opportunity to shape the SRE function, reduce operational toil and establish new ways of working across a large technology organisation.

The position is remote-first, with an expectation of working from the London office approximately one day per week. The salary is £90k plus cash benefits and bonus.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (LON)
Senior Site Reliability Engineer (LON)

McNally Recruitment Ltd • Greater London

Hybrid
GBP 90,000 - 150,000
Benefits as Cash
Hybrid work model
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Pathfinder • City Of London

Hybrid
GBP 81,000 - 99,000
Senior SRE (AWS)
Senior SRE (AWS)

VIQU IT Recruitment • Kingston

On-site
GBP 68,000 - 83,000
Bonus
On-call allowance
SRE Technical Lead
SRE Technical Lead

83zero Ltd • United Kingdom

Hybrid
GBP 90,000 - 110,000
Salary up to 100,000
5% annual bonus
Hybrid working model
+1
Site Reliability Engineer
Site Reliability Engineer

SR2 | Socially Responsible Recruitment | Certified B Corporation • Slough

On-site
GBP 65,000 - 90,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000
DevOps / Site Reliability Engineer (SRE)
DevOps / Site Reliability Engineer (SRE)

SCC • United Kingdom

Hybrid
GBP 70,000 - 110,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

National Health Service • Greater London

Hybrid
GBP 42,000 - 52,000
Hybrid working
Flexible working arrangements
On-site core HQs
Site Reliability Engineer
Site Reliability Engineer

Biometric Talent Ltd • Manchester

On-site
GBP 40,000 - 65,000
Performance-Based Bonus
Pension Scheme
Hybrid Working
+2
Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

On-site
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts