Operations and SRE Manager

LexisNexis Risk Solutions

Sutton

Hybrid

GBP 90,000 - 130,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

ICIS is seeking an SRE Manager to lead the reliability function for production services and guide AI-Ops initiatives. You will drive observability, incident response, and continuous improvement across internal and external customers.

The role requires strong people leadership, hands-on incident involvement, and the ability to align Ops with product priorities. You will inspire a team focused on resilience, automation, and proactive problem solving, delivering reliable services at scale with

Qualifications

  • Strong people leadership and coaching of team leads and engineers.
  • Practical experience with SRE, incident management, problem management and RCAs.
  • Experience in disaster recovery and operational resilience.
  • Ability to influence across Ops, engineering, product and wider tech teams.
  • Strong use of metrics, SLIs/SLOs, observability data to improve reliability.
  • Identifying automation opportunities and advancing AI-Ops and proactive operations.
  • Clear communication of risks, trade-offs, and needs up the chain.
  • Willingness to be hands-on during on-call and major incidents.

Responsibilities

  • Lead the team’s strategic direction and translate priorities into operational plans and backlogs.
  • Line manage and develop team leads, setting expectations and coaching.
  • Ensure incident management, problem management, RCAs, post-mortems are owned and driven to completion.
  • Strengthen operational processes and delegation effectiveness.
  • Drive SRE practices across observability, automation, disaster recovery and reliability design.
  • Protect service levels balancing engineering effort with roadmap and maintenance.
  • Coordinate with Ops, other tech functions, product teams and stakeholders on priorities and risks.

Skills

People leadership
SRE
Incident management
Post-mortems
RCA quality
Disaster recovery
Observability
Automation
AI-Ops
On-call readiness

Job description

## Operations and SRE ManagerApply: UK - Sutton (Carshalton): London Wall: Full time: Posted Today: R117696**Are you passionate about improving reliability at scale and helping teams deliver resilient, high-performing services?****Do you enjoy leading people, driving operational excellence, and shaping the future of AI-enabled operations?****About the Business:**At ICIS, our mission is to optimize the world's resources. We help companies make strategic, sustainable decisions by bringing transparency to markets across the world. We create a comprehensive view of commodities markets, providing companies with the data and intelligence to successfully navigate across global value chains every day. Our customers benefit from instant access to price assessments, reports and forecasts, a dedicated news channel, and supply and demand data. You can learn more about ICIS at https://www.icis.com/explore.**About our Team:**Our teams are fuelled by curiosity, relentlessly pursuing better customer outcomes. We're on a mission to deliver an unparalleled customer experience, excelling in communication. In a fast-paced environment, we thrive, embracing change with flexibility and composure under pressure. Our high-energy, self-motivated individuals are driven by a genuine desire to make a positive mark on our business. But that's not all - we're creative problem solvers with an entrepreneurial spirit.**About the Role:**As SRE Manager, you will lead the operational reliability function for ICIS, ensuring stable, resilient, and well-supported production services for both internal and external customers. You will be responsible for driving reliability improvements, advancing automation and AI-Ops capabilities, and leading a team focused on observability, incident response, operational excellence, and continuous improvement.**Responsibilities:*** Lead the implementation of the team’s strategic direction, translating priorities into clear operational plans, backlogs and deliverables.* Line manage and develop the team leaders, setting clear expectations, providing coaching, challenging unhelpful patterns and helping them build stronger delivery discipline.* Ensure incident management, problem management, RCAs, post-mortems and improvement actions are owned, tracked and completed.* Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective.* Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support.* Protect service levels by ensuring engineering effort is balanced across InfoSec commitments, operational tickets, roadmap delivery and mid-term maintenance.* Act as the connection point between Ops, wider technology functions, product teams and business stakeholders, ensuring priorities, risks and trade-offs are clearly communicated.* Help shape the team for AI-Ops by identifying practical automation, alerting, triage and remediation opportunities.**Requirements*** Strong people leadership, including coaching, performance management, prioritisation and development of team leads and engineers.* Practical experience of SRE, incident management, problem management, post-mortems, RCA quality, disaster recovery and operational resilience.* Ability to influence across Ops, engineering, product and wider technology teams, especially where priorities are competing or ownership is unclear.* Strong delivery discipline, with the ability to focus the team on fewer, higher-value initiatives and see them through to completion.* Experience using metrics, service levels, observability and operational data to drive better decisions and reliability improvements.* Ability to identify automation opportunities and help move the team towards AI-Ops, reduced toil and more proactive operations.* Clear communication style, able to explain operational risks, trade-offs, successes and support needs up and down the organisation.* Willingness to be hands-on when required, including participation in on-call and support of major incidents.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Operations and SRE Manager
Operations and SRE Manager

RELX Group • Carshalton

On-site
GBP 90,000 - 120,000
Operations and SRE Manager
Operations and SRE Manager

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
SRE & Reliability Lead — AI-Ops & Observability
SRE & Reliability Lead — AI-Ops & Observability

LexisNexis Risk Solutions • Carshalton

On-site
GBP 90,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

Falconsmartit • Hove

Hybrid
GBP 90,000 - 130,000
SRE & Operations Manager — AI-Driven Reliability
SRE & Operations Manager — AI-Driven Reliability

LexisNexis Risk Solutions • Sutton

Hybrid
GBP 90,000 - 130,000
Observability SRE
Observability SRE

HCLTech • Greater London

On-site
GBP 70,000 - 95,000
SRE Architect
SRE Architect

Hitachi • Greater London

On-site
GBP 42,000 - 70,000
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom
SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hitachids • Greater London

On-site
GBP 90,000 - 140,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000