Operations and SRE Manager

LexisNexis Risk Solutions

United Kingdom

Remote

GBP 90,000 - 130,000

Full time

6 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

ICIS is seeking an SRE Manager to lead the reliability function for production services used by internal and external customers. You will drive reliability improvements, advance automation and AI-Ops capabilities, and lead a team focused on observability, incident response, and continuous improvement.

You will line manage team leaders, shape strategic direction, ensure incident management and RCAs are completed, and collaborate with Ops, engineering, product and stakeholders to balance

Qualifications

  • Leadership of SRE teams and engineers
  • Practical SRE experience incl. incident and disaster recovery
  • Ability to influence across Ops, Eng, Product and other teams
  • Delivery focus with prioritised high-value initiatives
  • Experience using SLIs/SLOs and observability data
  • Identification of automation opportunities for AI-Ops
  • Clear, concise communication of risks and needs
  • Willingness to be hands-on during major incidents

Responsibilities

  • Lead the implementation of the team’s strategic direction with clear plans and backlogs.
  • Line manage and develop team leaders, set expectations and coach for delivery discipline.
  • Ensure incident management, RCAs and post-mortems are owned and closed.
  • Strengthen processes with clear ownership and effective delegation.
  • Drive SRE practices across observability, automation, DR and on-call readiness.
  • Balance engineering work with InfoSec, tickets, roadmap and maintenance.
  • Connect Ops with product teams and business stakeholders on priorities.
  • Identify practical automation opportunities to move toward AI-Ops.

Skills

People leadership
SRE experience
Cross-functional influence
Delivery discipline
Observability metrics
Automation opportunities
Clear communication
Hands-on on-call

Job description

**Are you passionate about improving reliability at scale and helping teams deliver resilient, high-performing services?****Do you enjoy leading people, driving operational excellence, and shaping the future of AI-enabled operations?****About the Business:**At ICIS, our mission is to optimize the world's resources. We help companies make strategic, sustainable decisions by bringing transparency to markets across the world. We create a comprehensive view of commodities markets, providing companies with the data and intelligence to successfully navigate across global value chains every day. Our customers benefit from instant access to price assessments, reports and forecasts, a dedicated news channel, and supply and demand data. You can learn more about ICIS at https://www.icis.com/explore.**About our Team:**Our teams are fuelled by curiosity, relentlessly pursuing better customer outcomes. We're on a mission to deliver an unparalleled customer experience, excelling in communication. In a fast-paced environment, we thrive, embracing change with flexibility and composure under pressure. Our high-energy, self-motivated individuals are driven by a genuine desire to make a positive mark on our business. But that's not all - we're creative problem solvers with an entrepreneurial spirit.**About the Role:**As SRE Manager, you will lead the operational reliability function for ICIS, ensuring stable, resilient, and well-supported production services for both internal and external customers. You will be responsible for driving reliability improvements, advancing automation and AI-Ops capabilities, and leading a team focused on observability, incident response, operational excellence, and continuous improvement.**Responsibilities:*** Lead the implementation of the team’s strategic direction, translating priorities into clear operational plans, backlogs and deliverables.* Line manage and develop the team leaders, setting clear expectations, providing coaching, challenging unhelpful patterns and helping them build stronger delivery discipline.* Ensure incident management, problem management, RCAs, post-mortems and improvement actions are owned, tracked and completed.* Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective.* Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support.* Protect service levels by ensuring engineering effort is balanced across InfoSec commitments, operational tickets, roadmap delivery and mid-term maintenance.* Act as the connection point between Ops, wider technology functions, product teams and business stakeholders, ensuring priorities, risks and trade-offs are clearly communicated.* Help shape the team for AI-Ops by identifying practical automation, alerting, triage and remediation opportunities.**Requirements*** Strong people leadership, including coaching, performance management, prioritisation and development of team leads and engineers.* Practical experience of SRE, incident management, problem management, post-mortems, RCA quality, disaster recovery and operational resilience.* Ability to influence across Ops, engineering, product and wider technology teams, especially where priorities are competing or ownership is unclear.* Strong delivery discipline, with the ability to focus the team on fewer, higher-value initiatives and see them through to completion.* Experience using metrics, service levels, observability and operational data to drive better decisions and reliability improvements.* Ability to identify automation opportunities and help move the team towards AI-Ops, reduced toil and more proactive operations.* Clear communication style, able to explain operational risks, trade-offs, successes and support needs up and down the organisation.* Willingness to be hands-on when required, including participation in on-call and support of major incidents.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Sutton

On-site
GBP 90,000 - 130,000
Operations and SRE Manager
Operations and SRE Manager

RELX Group • Carshalton

On-site
GBP 90,000 - 120,000
Operations and SRE Manager
Operations and SRE Manager

Talent Octopusventures • City Of London

On-site
GBP 90,000 - 140,000
Country-specific benefits
SRE & Operations Manager — AI-Driven Reliability
SRE & Operations Manager — AI-Driven Reliability

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
SRE & Operations Leader - Reliability at Scale
SRE & Operations Leader - Reliability at Scale

LexisNexis Risk Solutions • United Kingdom

Remote
GBP 90,000 - 130,000
Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
SRE & Reliability Manager — AI-Ops, Observability
SRE & Reliability Manager — AI-Ops, Observability

RELX Group • Carshalton

On-site
GBP 90,000 - 120,000
SRE & Reliability Lead — AI-Ops & Observability
SRE & Reliability Lead — AI-Ops & Observability

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
Observability SRE
Observability SRE

HCLTech • Greater London

On-site
GBP 70,000 - 95,000
Service Manager, Site Reliability Engineering
Service Manager, Site Reliability Engineering

Allstate Northern Ireland Limited • Belfast City District

Hybrid
GBP 75,000 - 95,000