Senior SRE

Pulse Recruit

Greater London

Hybrid

GBP 65,000 - 85,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

A technology organisation in London is seeking a Senior Site Reliability Engineer (SRE) to enhance and secure the reliability of systems for millions of users. You will design automated monitoring systems, manage SLOs, and collaborate with various teams to foster a reliable cloud-native platform. Ideal candidates will have a strong SRE or DevOps background, hands-on cloud experience, and a passion for automation. This hybrid role offers significant engineering challenges in a mission-driven environment.

Qualifications

  • Strong background in SRE, DevOps or Platform Engineering.
  • Experience running production systems at scale.
  • Understanding of observability, monitoring, and reliability.

Responsibilities

  • Design and automate monitoring and observability systems.
  • Define and manage SLOs, SLIs, and error budgets.
  • Support incident response and post-mortems.

Skills

SRE, DevOps or Platform Engineering
Automation
Reliability principles
Cloud infrastructure
Kubernetes
Complex systems debugging

Tools

GCP
AWS
Terraform
Python
Go
Prometheus
Grafana
Datadog

Job description

Location: London (Hybrid – 1 day per week in office)

We are working with a mission-led technology organisation that is continuing to scale a fully cloud-native platform as part of a major initiative. As they move away from traditional data centres, they are investing heavily in building a highly reliable, scalable and observable cloud platform.

As a Senior SRE, you will play a key role in ensuring the reliability and performance of systems that support millions of customers. This is a hands‑on engineering role where you will work closely with platform, cloud and product teams to embed reliability into everything they build.

You will be solving complex engineering problems across distributed systems, helping improve observability, automation and incident response as the platform continues to scale.

Key Responsibilities
  • Designing, improving and automating monitoring and observability systems
  • Defining and managing SLOs, SLIs and error budgets
  • Supporting incident response, root cause analysis and post-mortems
  • Working with engineering teams to design resilient, fault‑tolerant systems
  • Driving automation across infrastructure, deployments and operations
  • Contributing to capacity planning, performance tuning and cost optimisation
  • Participating in design reviews to improve reliability and scalability
Tech Environment
  • GCP and AWS
  • Kubernetes and containerised workloads
  • Terraform and Infrastructure as Code
  • Prometheus, Grafana, Datadog and modern observability tooling
  • CI/CD pipelines and automation tooling
  • Python, Go or similar scripting languages
  • Distributed systems at scale
About You
  • Strong background in SRE, DevOps or Platform Engineering
  • Experience running and supporting production systems at scale
  • Strong understanding of observability, monitoring and reliability principles
  • Hands‑on experience with cloud infrastructure and Kubernetes
  • Experience with Infrastructure as Code (Terraform or similar)
  • Comfortable debugging complex systems across infrastructure and application layers
  • Passionate about automation and improving engineering efficiency

This is a great opportunity to join a team building a platform with real‑world impact, combining complex engineering challenges with a mission to contribute to a future.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 70,000 - 90,000
Healthcare
Retirement planning
Paid volunteering days
+1
SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
Senior Site Reliability Engineer (Application / API Focused)
Senior Site Reliability Engineer (Application / API Focused)

Xpertise Recruitment • Greater London

Hybrid
GBP 80,000 - 100,000
25% Bonus
Excellent Benefits
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000
Site Reliability Engineer
Site Reliability Engineer

ReVybe IT Recruitment Limited • Greater London

Hybrid
GBP 51,000 - 85,000
Bonus
Benefits
Senior Platform Engineer / SRE
Senior Platform Engineer / SRE

Myn • Greater London

Hybrid
GBP 90,000 - 120,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GCS Recruitment • Knutsford

Hybrid
GBP 85,000 - 120,000
SRE Architect (68019)
SRE Architect (68019)

Hitachi Digital Services • Greater London

On-site
GBP 90,000 - 150,000
SRE
SRE

Source Group International • Greater London

Hybrid
GBP 68,000 - 108,000
SRE / DevOps Engineers
SRE / DevOps Engineers

HCLTech • Greater London

On-site
GBP 60,000 - 80,000