Site Reliability Engineer - NS London

BAE Systems Digital Intelligence

Greater London

Hybrid

GBP 50,000 - 70,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Hybrid working environment
On-call allowances
Overtime benefits for night shifts

Job summary

BAE Systems Digital Intelligence in Greater London is seeking a Site Reliability Engineer. In this role, you'll combine operational support and software engineering to automate processes, improve system reliability, and support essential services for a national security customer. You will be involved in a 24/7 on-call duty and participate in a progressive hybrid working environment that values flexibility. Ideal candidates will have software development experience and knowledge of database technologies, as well as monitoring large systems to enhance performance and stability.

Qualifications

  • Proven software development experience in web technologies and object-oriented programming.
  • Strong knowledge of Oracle SQL, MongoDB, PostgreSQL.
  • Skilled in Linux and Windows command line operations.
  • Experience with monitoring tools like Grafana, Prometheus, ELK, Splunk.
  • Capable of diagnosing application issues that lead to service outages.

Responsibilities

  • Support and maintain essential services for core mission applications.
  • Participate in a 24/7 on-call rota for critical production systems.
  • Automate manual operations to enhance efficiency.
  • Advise development teams on design and building practices.
  • Design and deploy monitoring tools for system observability.

Skills

Software development experience in web technologies
Knowledge of database technologies
Comfortable with Linux and Windows command lines
Experience monitoring large systems
Experience working in Agile teams
Diagnosing and troubleshooting application issues
Cross‑stack troubleshooting skills
Understanding of ITIL
Experience with micro‑services and Docker
Awareness of emerging technology trends

Job description

Location(s): [[mfield3]]

Site Reliability Engineering is a rapidly growing concept in industry, with a remit to drive the quality, reliability and performance of essential systems. As a Site Reliability Engineer you will be part of a team in BAE Systems at the forefront of this, delivering benefits to a key national security customer. We are building our team and tools, and will create a culture of continual improvement to revolutionise how our customer’s systems are built and maintained. This role blends operational product support with software engineering to create applications to understand overall system health. The SRE team sits within a wider programme at the core of the customer mission.

Role Holder

As an SRE, you will perform tasks historically done by operations teams, but using software and systems engineering expertise to replace manual labour with automation. The goal is to limit manual operations such as incident tickets and on‑call duties to no more than half of the team's time (and preferably less). You should have enthusiasm to learn and experiment, develop tools for application health, and improve reliability to support the customer mission.

Responsibilities
  • Support and maintain essential services that support core mission applications, proactively enhancing their availability, performance and stability.
  • Participate in the 24/7 on‑call rota, supporting critical production systems out of business hours; additional on‑call allowances and overtime benefits will be paid.
  • Find innovative solutions to problems rather than repeatable work, automating everything possible.
  • Work alongside development teams, advising them on good practices for designing and building systems.
  • Design and deploy monitoring products, creating bespoke tools where required, to provide comprehensive and intelligent observations that meet customer requirements and demonstrate daily improvements.
  • Be well‑versed in the relationship between software and infrastructure, understanding characteristics that enable scalability and resilience.
  • Participate in the wider DevOps/SRE community within the organisation.
Qualifications
  • Software development experience in web technologies and object‑oriented programming.
  • Knowledge of database technologies such as Oracle SQL, MongoDB, PostgreSQL.
  • Comfortable with Linux and Windows command lines (e.g. Bash, PowerShell).
  • Experience monitoring large systems using Grafana, Prometheus, ELK, Splunk.
  • Experience working in Agile teams and related tooling (e.g. Atlassian).
  • Diagnosing and troubleshooting application issues that result in service outages.
  • Cross‑stack troubleshooting skills.
  • Understanding of ITIL.
  • Experience with micro‑services, Docker, and container platforms such as OpenShift, Kubernetes.
  • Awareness of emerging technology trends to adopt cutting‑edge tools.
Security Clearance

Successful candidates must hold an active eDV clearance before applying.

Benefits & Work Environment

We embrace hybrid working, allowing flexibility in location and schedule to balance work and personal life. On‑call allowances and overtime benefits are provided for night shifts.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

Hybrid
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits
Site Reliability Engineer
Site Reliability Engineer

Anson McCade • Manchester

On-site
GBP 70,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

Anson McCade • Gloucester

On-site
GBP 60,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
Site Reliability Engineer
Site Reliability Engineer

Sanderson Government & Defence • Greater London

Hybrid
GBP 35,000 - 75,000
Flexible salary range reflecting seniority
Hybrid working and startup autonomy
Opportunity to shape platforms and engineering practices
Director of Site Reliability Engineering
Director of Site Reliability Engineering

EPAM Systems • Greater London

Hybrid
GBP 180,000 - 240,000
ESPP
Life Assurance
Income protection
+14
Site Reliability Engineer - Gloucester
Site Reliability Engineer - Gloucester

Hackajob Ltd • Gloucester

Hybrid
GBP 59,000 - 72,000
SRE Architect
SRE Architect

Hitachi • Greater London

On-site
GBP 42,000 - 70,000
Site Reliability Engineer
Site Reliability Engineer

SR2 | Socially Responsible Recruitment | Certified B Corporation • Slough

Hybrid
GBP 65,000 - 90,000