Site Reliability Engineer – NS London

BAE Systems

Greater London

Hybrid

GBP 45,000 - 70,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Hybrid working flexibility
On-call allowances
Overtime benefits

Job summary

BAE Systems is looking for a Site Reliability Engineer to enhance system performance, reliability, and observability for critical applications. You will automate manual operations, participate in a 24/7 on-call rota, and work in a collaborative Agile environment. Ideal candidates will have software development experience, knowledge of database technologies, and familiarity with system monitoring tools. BAE Systems promotes hybrid working to balance personal and professional life.

Qualifications

  • At least 3 years of software development experience.
  • Familiarity with database technologies and system monitoring tools.
  • Experience with Linux and Windows environments.

Responsibilities

  • Support and maintain essential services for core mission applications.
  • Automate processes to reduce manual operations.
  • Design and deploy monitoring products to enhance system observability.
  • Collaborate with development teams to improve system design.

Skills

Software development experience in web technologies
Object-oriented programming
Database technologies (Oracle SQL, MongoDB, PostgreSQL)
Linux and Windows command lines
Experience monitoring systems (Grafana, Prometheus, ELK, Splunk)
Agile methodologies
Diagnosing application issues
Cross-stack troubleshooting skills
Understanding of ITIL
Micro-services, Docker, Kubernetes

Job description

Location(s): [[mfield3]]

Site Reliability Engineering is a rapidly growing concept in industry, with a remit to drive the quality, reliability and performance of essential systems. As a Site Reliability Engineer you will be part of a team in BAE Systems at the forefront of this, delivering benefits to a key national security customer. We are building our team and tools, and will create a culture of continual improvement to revolutionise how our customer’s systems are built and maintained. This role blends operational product support with software engineering to create applications to understand overall system health. The SRE team sits within a wider programme at the core of the customer mission.

Role Holder

As an SRE, you will perform tasks historically done by operations teams, but using software and systems engineering expertise to replace manual labour with automation. The goal is to limit manual operations such as incident tickets and on‑call duties to no more than half of the team's time (and preferably less). You should have enthusiasm to learn and experiment, develop tools for application health, and improve reliability to support the customer mission.

Responsibilities
  • Support and maintain essential services that support core mission applications, proactively enhancing their availability, performance and stability.
  • Participate in the 24/7 on‑call rota, supporting critical production systems out of business hours; additional on‑call allowances and overtime benefits will be paid.
  • Find innovative solutions to problems rather than repeatable work, automating everything possible.
  • Work alongside development teams, advising them on good practices for designing and building systems.
  • Design and deploy monitoring products, creating bespoke tools where required, to provide comprehensive and intelligent observations that meet customer requirements and demonstrate daily improvements.
  • Be well‑versed in the relationship between software and infrastructure, understanding characteristics that enable scalability and resilience.
  • Participate in the wider DevOps/SRE community within the organisation.
Qualifications
  • Software development experience in web technologies and object‑oriented programming.
  • Knowledge of database technologies such as Oracle SQL, MongoDB, PostgreSQL.
  • Comfortable with Linux and Windows command lines (e.g. Bash, PowerShell).
  • Experience monitoring large systems using Grafana, Prometheus, ELK, Splunk.
  • Experience working in Agile teams and related tooling (e.g. Atlassian).
  • Diagnosing and troubleshooting application issues that result in service outages.
  • Cross‑stack troubleshooting skills.
  • Understanding of ITIL.
  • Experience with micro‑services, Docker, and container platforms such as OpenShift, Kubernetes.
  • Awareness of emerging technology trends to adopt cutting‑edge tools.
Security Clearance

Successful candidates must hold an active eDV clearance before applying.

Benefits & Work Environment

We embrace hybrid working, allowing flexibility in location and schedule to balance work and personal life. On‑call allowances and overtime benefits are provided for night shifts.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

Hybrid
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts
Site Reliability Engineer
Site Reliability Engineer

Anson McCade • Manchester

On-site
GBP 70,000 - 90,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer
Site Reliability Engineer

Sanderson Government & Defence • Greater London

Hybrid
GBP 35,000 - 75,000
Flexible salary range reflecting seniority
Hybrid working and startup autonomy
Opportunity to shape platforms and engineering practices
Site Reliability Engineer
Site Reliability Engineer

慨正橡扯 • Manchester

Hybrid
GBP 60,000 - 80,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 70,000 - 90,000
Healthcare
Retirement planning
Paid volunteering days
+1
Site Reliability Engineer - Gloucester
Site Reliability Engineer - Gloucester

Hackajob Ltd • Gloucester

Hybrid
GBP 59,000 - 72,000
SRE Architect
SRE Architect

Hitachi • Greater London

On-site
GBP 42,000 - 70,000
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000