SRE & Reliability Manager — AI-Ops, Observability

RELX Group

Carshalton

On-site

GBP 90,000 - 120,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

ICIS is seeking an SRE Manager to lead the reliability function for production services. You will drive reliability improvements, automation and AI-Ops, and manage a team focused on observability, incident response and continuous improvement.

You will translate priorities into operational plans, line-manage team leaders, and ensure RCAs, post-mortems and remediation actions are completed, while balancing workload with InfoSec commitments and product roadmaps.

Qualifications

  • Strong people leadership, including coaching, performance management and development of team leads and engineers.
  • Practical experience of SRE, incident management, post-mortems and RCA quality.
  • Experience delivering observability, automation and disaster recovery improvements.

Responsibilities

  • Lead the implementation of the team's strategic direction and translate priorities into operational plans and backlogs.
  • Line manage and develop team leaders, setting expectations and coaching for delivery discipline.
  • Ensure incident management, problem management, RCAs and post-mortems are owned and completed.
  • Strengthen operational process adherence with clear responsibilities and effective delegation.
  • Drive SRE practices across observability, automation and on-call readiness for production support.
  • Balance engineering efforts between security, tickets, roadmap and mid-term maintenance.
  • Communicate priorities, risks and trade-offs across Ops, engineering, product and business stakeholders.

Skills

People leadership
SRE
Incident management
Observability
Automation
AI-Ops

Job description

ICIS is seeking an SRE Manager to lead the reliability function for production services. You will drive reliability improvements, automation and AI-Ops, and manage a team focused on observability, incident response and continuous improvement.

You will translate priorities into operational plans, line-manage team leaders, and ensure RCAs, post-mortems and remediation actions are completed, while balancing workload with InfoSec commitments and product roadmaps.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE & Operations Manager — AI-Driven Reliability
SRE & Operations Manager — AI-Driven Reliability

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
SRE & Reliability Lead — AI-Ops & Observability
SRE & Reliability Lead — AI-Ops & Observability

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
Operations and SRE Manager
Operations and SRE Manager

RELX Group • Carshalton

On-site
GBP 90,000 - 120,000
SRE Architect: Reliability, Observability & Automation Lead
SRE Architect: Reliability, Observability & Automation Lead

Hitachi Digital Services • Greater London

On-site
GBP 90,000 - 150,000
Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
SRE Architect
SRE Architect

Hitachi • Greater London

On-site
GBP 42,000 - 70,000
SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
Senior SRE, Observability & Cloud Reliability
Senior SRE, Observability & Cloud Reliability

United States Digital Space LLC • Greater London

Hybrid
GBP 120,000 - 170,000
Hybrid work up to 3 days per week
Senior SRE: Reliability, Observability & Automation Lead (Hybrid)
Senior SRE: Reliability, Observability & Automation Lead (Hybrid)

GCS Recruitment • Knutsford

Hybrid
GBP 85,000 - 120,000
SRE for AI-Driven Financial Infrastructure
SRE for AI-Driven Financial Infrastructure

United States Digital Space LLC • Greater London

On-site
GBP 90,000 - 130,000
Daily catered lunches
Modern office environment
Tech talks and knowledge sharing