Senior Site Reliability Engineer

Source Technology Limited

Colchester

On-site

GBP 85,000 - 110,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Source Group International Ltd in the United Kingdom is seeking a Senior Site Reliability Engineer to enhance reliability, availability, and scalability of our SaaS platform. You will blend software engineering with cloud infrastructure, automation, and operational leadership to deliver resilient systems that enable rapid product delivery.

You will own production reliability, define SLI/SLO, lead blameless postmortems, and automate toil.

Qualifications

  • Proven experience as a Senior/experienced SRE with software engineering background.
  • Strong knowledge of SRE principles, incident management and observability.
  • Hands-on with Azure, Kubernetes, IaC and monitoring platforms.

Responsibilities

  • Own the reliability and performance of production services.
  • Define and operate SLIs, SLOs, and error budgets with engineering teams.
  • Lead incident response and drive blameless postmortems and continuous improvement.
  • Automate operational processes and reduce manual toil through engineering.
  • Build and operate cloud-native platforms using Azure, Kubernetes, and Infrastructure as Code.
  • Develop observability through effective monitoring, alerting, and telemetry.
  • Mentor engineers and promote reliability best practices across the organisation.

Skills

SRE principles
Automation
Observability
Incident management
Cloud platforms
Azure
Kubernetes
Infrastructure as Code
Monitoring
Software engineering

Tools

Prometheus
Grafana
Datadog

Job description

Overview

The Senior Site Reliability Engineer is responsible for the reliability, availability, scalability, and operational excellence of our SaaS platform. This role combines software engineering, cloud infrastructure, automation, and operational leadership to build resilient systems that enable rapid product delivery.

Responsibilities
  • Own the reliability and performance of production services.
  • Define and operate SLIs, SLOs, and error budgets with engineering teams.
  • Lead incident response and drive blameless postmortems and continuous improvement.
  • Automate operational processes and reduce manual toil through engineering.
  • Build and operate cloud-native platforms using Azure, Kubernetes, and Infrastructure as Code.
  • Develop observability through effective monitoring, alerting, and telemetry.
  • Mentor engineers and promote reliability best practices across the organisation.
Experience
  • Proven experience as a Senior or experienced Site Reliability Engineer with a software engineering background.
  • Ability to diagnose and make safe changes to PHP and Java or .NET applications.
  • Experience operating large-scale production SaaS systems.
  • Strong knowledge of SRE principles, incident management, observability, and operational excellence.
  • Hands‑on experience with Azure, Kubernetes, Infrastructure as Code, and monitoring platforms such as Prometheus, Grafana, or Datadog.
  • Experience influencing engineering teams and driving reliability improvements through collaboration and technical leadership.
Key Attributes
  • Passion for building reliable production systems.
  • Software engineering mindset with a focus on automation.
  • Strong technical judgement and ownership.
  • Excellent communication during incidents and day‑to‑day collaboration.
  • Commitment to blameless culture and continuous improvement.

By applying for this role, you consent to SGI contacting you by telephone regarding recruitment services, market updates, and relevant business opportunities.

We believe in equal opportunity for all and actively encourage applications from diverse backgrounds, experiences, and perspectives.

Source Group International Ltd is acting as an Employment Business in relation to this vacancy.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Reward Gateway • Greater London

Hybrid
GBP 70,000 - 110,000
Life assurance
Pension
Employee Share Plan
+3
Senior Site Reliability Engineer - Cloud-Native Leader
Senior Site Reliability Engineer - Cloud-Native Leader

Source Technology Limited • Colchester

On-site
GBP 85,000 - 110,000
Head of Site Reliability Engineering – SRE
Head of Site Reliability Engineering – SRE

Jobtailor • Bristol

On-site
GBP 110,000 - 140,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LSEG • Nottingham

On-site
GBP 70,000 - 90,000
Healthcare
Retirement planning
Paid volunteering days
+1
Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Senior Site Reliability Engineer (Application / API Focused)
Senior Site Reliability Engineer (Application / API Focused)

Xpertise Recruitment • Greater London

Hybrid
GBP 80,000 - 100,000
25% Bonus
Excellent Benefits
SRE Architect (68019)
SRE Architect (68019)

Hitachi Digital Services • Greater London

On-site
GBP 90,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

ScaleneWorks People Solutions LLP • Bournemouth

On-site
GBP 60,000 - 80,000
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000