Operations and SRE Manager

RELX Group

Carshalton

On-site

GBP 90,000 - 120,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

ICIS is seeking an SRE Manager to lead the reliability function for production services. You will drive reliability improvements, automation and AI-Ops, and manage a team focused on observability, incident response and continuous improvement.

You will translate priorities into operational plans, line-manage team leaders, and ensure RCAs, post-mortems and remediation actions are completed, while balancing workload with InfoSec commitments and product roadmaps.

Qualifications

  • Strong people leadership, including coaching, performance management and development of team leads and engineers.
  • Practical experience of SRE, incident management, post-mortems and RCA quality.
  • Experience delivering observability, automation and disaster recovery improvements.

Responsibilities

  • Lead the implementation of the team's strategic direction and translate priorities into operational plans and backlogs.
  • Line manage and develop team leaders, setting expectations and coaching for delivery discipline.
  • Ensure incident management, problem management, RCAs and post-mortems are owned and completed.
  • Strengthen operational process adherence with clear responsibilities and effective delegation.
  • Drive SRE practices across observability, automation and on-call readiness for production support.
  • Balance engineering efforts between security, tickets, roadmap and mid-term maintenance.
  • Communicate priorities, risks and trade-offs across Ops, engineering, product and business stakeholders.

Skills

People leadership
SRE
Incident management
Observability
Automation
AI-Ops

Job description

As SRE Manager, you will lead the operational reliability function for ICIS, ensuring stable, resilient, and well-supported production services for both internal and external customers. You will be responsible for driving reliability improvements, advancing automation and AI-Ops capabilities, and leading a team focused on observability, incident response, operational excellence, and continuous improvement.,

  • Lead the implementation of the team's strategic direction, translating priorities into clear operational plans, backlogs and deliverables.
  • Line manage and develop the team leaders, setting clear expectations, providing coaching, challenging unhelpful patterns and helping them build stronger delivery discipline.
  • Ensure incident management, problem management, RCAs, post-mortems and improvement actions are owned, tracked and completed.
  • Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective.
  • Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support.
  • Protect service levels by ensuring engineering effort is balanced across InfoSec commitments, operational tickets, roadmap delivery and mid-term maintenance.
  • Act as the connection point between Ops, wider technology functions, product teams and business stakeholders, ensuring priorities, risks and trade-offs are clearly communicated.
  • Help shape the team for AI-Ops by identifying practical automation, alerting, triage and remediation opportunities.
    Our teams are fuelled by curiosity, relentlessly pursuing better customer outcomes. We're on a mission to deliver an unparalleled customer experience, excelling in communication. In a fast-paced environment, we thrive, embracing change with flexibility and composure under pressure. Our high-energy, self-motivated individuals are driven by a genuine desire to make a positive mark on our business. But that's not all - we're creative problem solvers with an entrepreneurial spirit., Strong people leadership, including coaching, performance management, prioritisation and development of team leads and engineers.
  • Practical experience of SRE, incident management, problem management, post-mortems, RCA quality, disaster recovery and operational resilience.
  • Ability to influence across Ops, engineering, product and wider technology teams, especially where priorities are competing or ownership is unclear.
  • Strong delivery discipline, with the ability to focus the team on fewer, higher-value initiatives and see them through to completion.
  • Experience using metrics, service levels, observability and operational data to drive better decisions and reliability improvements.
  • Ability to identify automation opportunities and help move the team towards AI-Ops, reduced toil and more proactive operations.
  • Clear communication style, able to explain operational risks, trade-offs, successes and support needs up and down the organisation.
  • Willingness to be hands-on when required, including participation in on-call and support of major incidents.
    Are you passionate about improving reliability at scale and helping teams deliver resilient, high-performing services?

Do you enjoy leading people, driving operational excellence, and shaping the future of AI-enabled operations? About the Business: At ICIS, our mission is to optimize the world's resources. We help companies make strategic, sustainable decisions by bringing transparency to markets across the world. We create a comprehensive view of commodities markets, providing companies with the data and intelligence to successfully navigate across global value chains every day. Our customers benefit from instant access to price assessments, reports and forecasts, a dedicated news channel, and supply and demand data. You can learn more about ICIS at https://www.icis.com/explore.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
SRE & Operations Manager — AI-Driven Reliability
SRE & Operations Manager — AI-Driven Reliability

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
SRE & Reliability Manager — AI-Ops, Observability
SRE & Reliability Manager — AI-Ops, Observability

RELX Group • Carshalton

On-site
GBP 90,000 - 120,000
SRE & Reliability Lead — AI-Ops & Observability
SRE & Reliability Lead — AI-Ops & Observability

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
SRE Architect
SRE Architect

Hitachi • Greater London

On-site
GBP 42,000 - 70,000
SRE Architect (68019)
SRE Architect (68019)

Hitachi Digital Services • Greater London

On-site
GBP 90,000 - 150,000
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000
Site Reliability Engineer
Site Reliability Engineer

Falconsmartit • Hove

Hybrid
GBP 90,000 - 130,000
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000