Hybrid SRE: Build Reliable, Scalable Systems

bet365

Manchester

Hybrid

GBP 70,000 - 110,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Hybrid work from home policy

Job summary

bet365 is seeking a Site Reliability Engineer to shape the stability of the systems behind every click, query and live change. The SRE team protects and improves availability, performance and resilience across a complex global product, using automation, observability and incident response.

You will work across SRE, development and IT Operations, embedding reliability in the software lifecycle and mentoring colleagues while leveraging Open Telemetry, Grafana, New Relic and Terraform to improve

Qualifications

  • Knowledge of modern development practices, including testing, source control and delivery lifecycles.
  • Understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management.
  • Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty.
  • Proficiency in shell scripting for automation and system management.
  • Experience with Infrastructure as Code, including Terraform and Ansible.
  • Knowledge of Cloudflare or a comparable edge platform, including DNS, CDN, WAF, DDoS protection and traffic management.
  • Ability to troubleshoot distributed systems across edge, network, platform, application, dependency and origin layers.
  • Experience in large-scale, 24/7 enterprise environments where uptime, performance and stability are critical.
  • Practical experience using LLM platforms and coding assistants safely to improve productivity and root-cause analysis.

Responsibilities

  • Develop and maintain resilient tools, operational APIs and automation for effective system management.
  • Use orchestration and scripting to remove manual activity, reduce toil and improve operational consistency.
  • Write and contribute to code, telemetry and instrumentation that improve service reliability and observability.
  • Build dashboards and operational views using telemetry from Grafana, Splunk, New Relic and related platforms.
  • Configure and manage Cloudflare edge services using Infrastructure as Code and integrate edge telemetry with observability platforms.
  • Diagnose incidents end to end, trace issues from the edge through to origin systems and coordinate remediation.
  • Participate in live incident response, post-mortems and root-cause analysis to prevent recurrence.
  • Maintain and administer monitoring, alerting, APM and analytics toolsets, including PagerDuty workflows.
  • Drive reliability, observability, and performance improvements across teams.
  • Mentor colleagues and collaborate with IT Operations to deliver tooling that increases business value.

Skills

SRE principles
Incident management
Automation
Shell scripting
Observability mindset
Team collaboration
Problem solving
LLM tooling usage

Tools

Terraform
Ansible
OpenTelemetry
Splunk
New Relic
Grafana
PagerDuty
Cloudflare

Job description

bet365 is seeking a Site Reliability Engineer to shape the stability of the systems behind every click, query and live change. The SRE team protects and improves availability, performance and resilience across a complex global product, using automation, observability and incident response.

You will work across SRE, development and IT Operations, embedding reliability in the software lifecycle and mentoring colleagues while leveraging Open Telemetry, Grafana, New Relic and Terraform to improve

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Hybrid, Observability
Site Reliability Engineer - Hybrid, Observability

bet365 Group • United Kingdom

Hybrid
GBP 75,000 - 110,000
Eye care and Flu Vaccinations
Life Assurance
Senior AI-Driven SRE & Software Engineer (Hybrid)
Senior AI-Driven SRE & Software Engineer (Hybrid)

bet365 • Manchester

Hybrid
GBP 70,000 - 110,000
Global Site Reliability Engineer (Hybrid)
Global Site Reliability Engineer (Hybrid)

bet365 • Burslem

Hybrid
GBP 70,000 - 110,000
Hybrid SRE: AI-Driven Reliability & Observability
Hybrid SRE: AI-Driven Reliability & Observability

bet365 Group • Manchester

Hybrid
GBP 60,000 - 80,000
Eye care
Flu vaccinations
Life assurance
SRE Lead: Scalable, Reliable Betting Platform (Hybrid)
SRE Lead: Scalable, Reliable Betting Platform (Hybrid)

evoke • Leeds

Hybrid
GBP 70,000 - 90,000
Family Support
Discounts at retailers
Competitive salary and bonus
+3
Software Engineer, SRE
Software Engineer, SRE

bet365 • Stoke-on-Trent

On-site
GBP 70,000 - 100,000
Hybrid working from home
SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
Site Reliability Engineer
Site Reliability Engineer

bet365 Group • United Kingdom

Hybrid
GBP 75,000 - 110,000
Eye care and Flu Vaccinations
Life Assurance
Site Reliability Engineer
Site Reliability Engineer

bet365 • Burslem

Hybrid
GBP 70,000 - 110,000
Site Reliability Engineer: AI-Powered Observability
Site Reliability Engineer: AI-Powered Observability

bet365 • Stoke-on-Trent

Hybrid
GBP 70,000 - 100,000
Hybrid working from home