Site Reliability Engineer

Anson McCade

Gloucester

On-site

GBP 60,000 - 90,000

Full time

33 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Anson McCade is seeking a Site Reliability Engineer in Gloucester to help keep critical applications available and performing at their best. You will work across software, infrastructure and operations with a strong emphasis on automation and continuous improvement.

The role involves collaborating with development and product teams, improving service design, monitoring, and resilience. You will automate tasks, drive reliability, and help standardise SRE/DevOps practices across the engineering

Qualifications

  • Background in SRE, DevOps, platform engineering, or similar infrastructure-focused role.
  • Experience with Linux and Windows, including Bash and PowerShell.
  • Experience with monitoring/logging tools (ELK/Elasticsearch).
  • Experience with Docker, containers and microservices.
  • Experience with deployment/configuration tools (Chef, Puppet).
  • Experience with Elasticsearch, MongoDB or similar databases.
  • Familiar with Jira, Agile/Scrum environments and ITIL terminology.

Responsibilities

  • Keep critical applications available, stable, and performing as required.
  • Collaborate with development and product teams to improve service design and support.
  • Automate repetitive operational tasks to reduce manual effort.
  • Enhance monitoring, logging, alerting, and overall service visibility.
  • Use data to identify and implement reliability and scalability improvements.
  • Contribute to SRE and DevOps standards across the engineering community.

Skills

SRE
Automation
Linux
Windows
Docker
Containers
Monitoring
Incident response

Tools

ELK/Elastic Stack
Kubernetes
Puppet
Chef
Jira

Job description

Security: Candidates must be eligible to achieve UKIC DV Clearance

I’m working on a Site Reliability Engineer opportunity with a growing national-security technology team in Gloucester.

The role sits across software engineering, infrastructure and operations, with a strong focus on automation and continuous improvement. Rather than simply responding to incidents and closing tickets, you’ll be expected to understand why issues are happening and put the right engineering solution in place.

The role

You’ll be responsible for keeping critical applications available, stable and performing as they should. That will involve working closely with development and product teams, improving the way services are designed and supported, and making sure systems are properly monitored and resilient.

The day-to-day work will include:
  • Supporting and improving services used by important mission applications.
  • Investigating incidents and resolving problems across the application and infrastructure stack.
  • Automating repetitive operational tasks wherever possible.
  • Improving monitoring, logging, alerting and overall service visibility.
  • Using performance and availability data to identify areas for improvement.
  • Advising development teams on reliability, scalability and supportability.
  • Working with cloud platforms, containers and microservices.
  • Contributing to SRE and DevOps standards across the wider engineering community.
  • Introducing practical improvements that reduce manual support and improve service quality.

This is not a role where you’ll spend all your time dealing with operational tickets. The expectation is that support work is kept under control so that the team can focus on automation, engineering improvements and making the underlying services more reliable.

What they’re looking for

You’ll ideally have a background in SRE, DevOps, platform engineering, software engineering or a similar infrastructure-focused role.

Experience with some of the following would be useful:

  • Linux and Windows, including Bash and PowerShell.
  • ELK, Elasticsearch or similar monitoring and logging tools.
  • Docker, containers and microservices.
  • Chef, Puppet or comparable deployment and configuration tools.
  • Elasticsearch, MongoDB or similar database technologies.
  • Application troubleshooting and incident resolution.
  • Agile or Scrum delivery environments and tools such as Jira.
  • ITIL terminology and service-management processes.
  • Automated testing tools such as Selenium.
  • Working with or improving open-source software.

You don’t need to tick every single box. A strong troubleshooting background, good software engineering principles and an ability to automate and improve systems will be more important than knowing every tool listed above.

The opportunity

This is a chance to work on technically challenging systems where availability, resilience and security genuinely matter. You’ll join an expanding engineering community in Gloucester and work alongside experienced software, infrastructure and DevOps professionals.

If your background is in SRE, DevOps, platform engineering or software-focused infrastructure and you’re looking for something more meaningful than standard operational support, get in touch with Chris Prendergast at Anson McCade.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Anson McCade • Manchester

On-site
GBP 70,000 - 90,000
Site Reliability Engineer - Gloucester
Site Reliability Engineer - Gloucester

Hackajob Ltd • Gloucester

Hybrid
GBP 59,000 - 72,000
Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

Hybrid
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts
Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

Hybrid
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits
Site Reliability Engineer
Site Reliability Engineer

Sanderson Government & Defence • Greater London

Hybrid
GBP 35,000 - 75,000
Flexible salary range reflecting seniority
Hybrid working and startup autonomy
Opportunity to shape platforms and engineering practices
Site Reliability Engineer - SC Cleared
Site Reliability Engineer - SC Cleared

Searchability NS&D • Gloucester

Hybrid
GBP 59,000 - 72,000
Security Engineer (Site Reliability Engineering) - SC Cleared
Security Engineer (Site Reliability Engineering) - SC Cleared

Sanderson Government & Defence • City Of London

Hybrid
GBP 125,000 - 136,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer (Edv) - National Security
Site Reliability Engineer (Edv) - National Security

Forward Role Recruitment • Cheltenham

On-site
GBP 65,000 - 90,000
Senior DevOps Engineer
Senior DevOps Engineer

Sopra Steria Ltd • Cheltenham

On-site
GBP 65,000 - 90,000
Car allowance
Annual leave
Private medical
+3