Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
RELX International in London is seeking a Senior Site Reliability Engineer to drive reliability, scalability and performance of core platforms. You will lead automation efforts and partner with engineering teams to reduce toil and improve recovery capabilities.
You will design observability, incident response and distributed systems improvements, advocating for high code quality and operational excellence across squads.
Are you passionate about building resilient, scalable systems that power mission-critical applications?
Do you thrive on automating operations, improving reliability, and ensuring exceptional system performance?
Embedded Innovation Teams are cross-functional squads embedded within our segments to rapidly turn internal AI experimentation into validated, reusable solutions, building the capabilities we need to deliver customer value and growth. We work problem-first rather than tool-first, directly inside segment and function teams, improving the internal workflows that help our people deliver better outcomes for customers, faster.
As a Senior Site Reliability Engineer (SRE), you will play a key role in ensuring the reliability, scalability, and performance of our critical platforms and services. You will lead complex reliability initiatives, drive automation efforts to reduce operational toil, and help build resilient systems that deliver exceptional customer experiences.
You will leverage your expertise in observability, incident response, and distributed systems to proactively identify and resolve reliability challenges. Working closely with engineering teams, you will design and implement solutions that improve service availability, streamline operations, and enhance system recovery capabilities.
You will hold a high bar on code quality, flag risks and blockers early, and work alongside host-function stakeholders to make sure what you build fits real workflows, not assumed ones. You will also support handover and capability-building so the solution is owned and operable after the squad moves on.