Senior Site Reliability Engineer

Selby Jennings

Greater London

On-site

GBP 70,000 - 90,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

A leading systematic hedge fund based in London is seeking a Senior Site Reliability Engineer to enhance the reliability and operability of their engineering platform. The ideal candidate will apply SRE principles, possess strong expertise in Linux, and have skills in Go and/or Python. Responsibilities include improving observability, building dashboards, and enhancing service reliability. Applicants should be technically proficient in Infrastructure-as-Code and container tools like Docker. This is an excellent opportunity for a highly skilled engineer looking to make a significant impact.

Qualifications

  • Hands-on experience applying SRE principles in production environments.
  • Strong technical expertise in Linux systems.
  • Experience building and operating containerized workloads.

Responsibilities

  • Own the effectiveness of the observability platform.
  • Build and maintain actionable dashboards and alerting.
  • Define and apply SLIs and SLOs for operational decision-making.

Skills

SRE principles application
Go programming
Python programming
Linux systems
Docker
PromQL
Infrastructure-as-Code
OpenTelemetry
Kubernetes

Job description

Our client, a leading systematic hedge fund, is seeking a Senior Site Reliability Engineer to join their London-based platform team. In this role, you will focus on enhancing the reliability, resilience, and day-to-day operability of a rapidly scaling engineering platform. You will work closely with software engineers and platform owners to strengthen observability, improve incident response processes, and drive measurable reliability outcomes.

To be successful, you will bring hands-on experience applying SRE principles in production environments, alongside strong expertise in Linux systems. You must be capable of building and operating containerized workloads using tools such as Docker or Podman, and hold strong experience in Go and/or Python. We are looking for a highly technical individual with strong Infrastructure-as-Code proficiency, and the ability to effectively query, interpret, and reason about metrics using PromQL. A key part of this role will involve owning and improving the overall effectiveness of the platform\'s observability.

Responsibilities
  • Own the effectiveness of the observability platform, ensuring high-quality signals, alert fidelity, and ongoing suitability as the platform scales.
  • Build and maintain actionable, low-noise dashboards and alerting across metrics and logs.
  • Define and apply SLIs and SLOs where they support operational decision-making.
  • Apply IaaC across observability and supporting systems.
  • Improve the reliability, scalability, and operability of core services through hands-on engineering changes.
Requirements
  • Strong practical experience applying SRE principles in production environments.
  • Strong development experience in Go and/or Python.
  • OpenTelemetry experience (metrics, logs, traces).
  • Kubernetes and cloud-native platform experience.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Selby Jennings • City Of London

On-site
GBP 90,000 - 140,000
Senior SRE: Observability & Platform Reliability
Senior SRE: Observability & Platform Reliability

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000
Senior Site Reliability Engineer - Selby Jennings
Senior Site Reliability Engineer - Selby Jennings

eFinancialCareers • Greater London

On-site
GBP 90,000 - 130,000
Senior Platform Engineer / SRE
Senior Platform Engineer / SRE

Myn • Greater London

Hybrid
GBP 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Tempest Vane Partners • Greater London

On-site
GBP 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

Selby Jennings • Greater London

On-site
GBP 80,000 - 100,000
Site Reliability Engineer - Tier-1 Quantitative Trading Company
Site Reliability Engineer - Tier-1 Quantitative Trading Company

Hamilton Barnes ? • Greater London

On-site
GBP 85,000 - 125,000
Site Reliability Engineer - Banking & Finance
Site Reliability Engineer - Banking & Finance

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 90,000 - 130,000
Global engineering organisation
Engineering-led culture
Technically challenging problems
+1
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

GCS • Glasgow

Hybrid
GBP 75,000 - 110,000