Senior SRE: Latency-Sensitive Platform + Equity

Understanding Recruitment

United Kingdom

On-site

GBP 90,000 - 120,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Understanding Recruitment is seeking a Senior Site Reliability Engineer to enhance reliability, observability, and tooling for a latency-sensitive production platform. The role spans production infrastructure, monitoring, incident response, and deployment workflows with a strong Linux and networking focus.

You will work to improve CI/CD, automation, and internal tooling while elevating developer experience from local setups to production.

Qualifications

  • Strong experience in Site Reliability Engineering, Platform Engineering, DevOps or Infrastructure Engineering.
  • Experience operating production infrastructure in cloud environments.
  • Strong Linux systems knowledge and understanding of networking fundamentals.
  • Experience with monitoring, observability and alerting.
  • Strong troubleshooting and root cause analysis skills.
  • Experience with CI/CD and infrastructure automation.
  • AWS, Terraform or Ansible experience would be advantageous.
  • Experience with high-performance, high-throughput or latency-sensitive systems would be valuable.

Responsibilities

  • Improve the reliability and operability of production systems.
  • Build and improve monitoring, logging, tracing, dashboards and alerting.
  • Improve incident diagnosis, root cause analysis and operational workflows.
  • Build safer and more repeatable deployment and rollback processes.
  • Automate repetitive operational and infrastructure work.
  • Improve CI/CD pipelines and release processes.
  • Develop internal tooling that helps engineers operate production systems more effectively.
  • Improve the developer experience from local development through to production.
  • Work with Linux systems, networking, host configuration and resource contention.
  • Contribute to infrastructure security, access controls, secrets management and system hardening.

Skills

Site reliability engineering
Platform engineering
DevOps
Infrastructure engineering
Cloud environments
Linux systems
Networking fundamentals
Monitoring & observability
CI/CD
Automation

Tools

AWS
Terraform
Ansible

Job description

Understanding Recruitment is seeking a Senior Site Reliability Engineer to enhance reliability, observability, and tooling for a latency-sensitive production platform. The role spans production infrastructure, monitoring, incident response, and deployment workflows with a strong Linux and networking focus.

You will work to improve CI/CD, automation, and internal tooling while elevating developer experience from local setups to production.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Understanding Recruitment • United Kingdom

On-site
GBP 90,000 - 120,000
Senior SRE / Production Engineer: High-Performance Linux
Senior SRE / Production Engineer: High-Performance Linux

Hunter Bond • Greater London

On-site
GBP 90,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000
Senior SRE: Observability & Platform Reliability
Senior SRE: Observability & Platform Reliability

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000
Senior DevOps Engineer: Latency Infra, Equity & Autonomy
Senior DevOps Engineer: Latency Infra, Equity & Autonomy

Understanding Recruitment • Greater London

On-site
GBP 90,000 - 120,000
Equity
Performance-based bonus
Senior Data Platform SRE — Hybrid, Obs & Reliability Leader
Senior Data Platform SRE — Hybrid, Obs & Reliability Leader

Outerlimit • Greater London

Hybrid
GBP 90,000 - 125,000
Hybrid work London
Conference attendance
Training budget
+1
Senior SRE: Build Resilient, AI-Driven Systems
Senior SRE: Build Resilient, AI-Driven Systems

Elsevier • City Of London

On-site
GBP 70,000 - 120,000
Senior SRE: Cloud Platform Reliability & Automation
Senior SRE: Cloud Platform Reliability & Automation

Selby Jennings • Greater London

On-site
GBP 90,000 - 130,000
Senior SRE: Architecting Reliability & Observability
Senior SRE: Architecting Reliability & Observability

London Stock Exchange Group • Nottingham

On-site
GBP 90,000 - 120,000
Healthcare
Retirement planning
Paid volunteering days
+1
Senior SRE: Cloud Reliability & Observability
Senior SRE: Cloud Reliability & Observability

Renesas Electronics Corp. • Cambridge

On-site
GBP 90,000 - 120,000