Site Reliability Engineer

London Stock Exchange Group

Creve Coeur (MO)

On-site

USD 120,000 - 160,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Healthcare
Retirement planning
Paid volunteering days
Wellbeing initiatives

Job summary

London Stock Exchange Group is hiring a Site Reliability Engineer to run and improve our internal observability platform. You will own reliability, supportability, and usability, shaping self-service patterns for teams to adopt consistently. This role focuses on the platform rather than application ownership.

You will enable trustful telemetry through dashboards, alerts, and service health insights, with a strong emphasis on automation, documentation, and scalable operations across environments.

Qualifications

  • Experience supporting production platforms in SRE or platform engineering roles.
  • Ability to use observability data to investigate issues.
  • Knowledge of incident response, problem management, and service readiness.

Responsibilities

  • Platform operations and reliability: monitor health, diagnose issues, support recovery across environments.
  • Contribute to dashboards, SLOs, SLIs, runbooks, resilience checks, and performance validation.
  • Drive incident response, root cause analysis and follow-up actions to reduce repeats.
  • Improve automation and GitOps-based workflows to reduce manual work.
  • Maintain documentation and self-service tooling for teams.

Skills

Observability
SRE practices
Cloud
Linux
CI/CD
Communication
Automation

Tools

OpenTelemetry
Grafana
Datadog
Cribl
ClickHouse
Redis
Flink

Job description

ABOUT US:

LSEG (London Stock Exchange Group) is more than a diversified global financial markets infrastructure and data business. We are dedicated, open-access partners with a commitment to excellence in delivering the services our customers expect from us. With extensive experience, deep knowledge and worldwide presence across financial markets, we enable businesses and economies around the world to fund innovation, handle risk and create jobs. It's how we've contributed to supporting the financial stability and growth of communities and economies globally for more than 300 years. Through a comprehensive suite of trusted financial market infrastructure services - and our open-access model - we provide the flexibility, stability and trust that enable our customers to pursue their ambitions with confidence and clarity. LSEG is headquartered in the United Kingdom, with significant operations in 70 countries across EMEA, North America, Latin America and Asia Pacific. We employ 25,000 people globally, more than half located in Asia Pacific. LSEG's ticker symbol is LSEG.

Our culture:

People are at the heart of what we do and drive the success of our business. Our culture of connecting, creating opportunity and delivering excellence shape how we think, how we do things and how we help our people fulfil their potential. We embrace diversity and actively seek to attract individuals with unique backgrounds and perspectives. We break down barriers and encourage teamwork, enabling innovation and rapid development of solutions that make a difference. Our workplace generates an enriching and exciting experience for our people and customers alike. Our vision is to build an inclusive culture in which everyone feels encouraged to fulfil their potential. We know that real personal growth cannot be achieved by simply climbing a career ladder – which is why we encourage and enable a wealth of avenues and interesting opportunities for everyone to broaden and deepen their skills and expertise. As a global organization spanning 70 countries and one rooted in a culture of growth, opportunity, diversity and innovation, LSEG is a place where everyone can grow, develop and fulfil your potential with relevant careers.

Role Profile:

We are hiring a Site Reliability Engineer to help us run and improve LSEG's internal observability platform - and we want you to be part of building something that matters! Our platform brings telemetry, dashboards, alerts, service health, and operational insight into one common experience. We use it to help engineering and service teams detect issues earlier, reduce incident impact, and improve operational reliability. The focus of this role is the platform itself - its reliability, supportability, and usability - and the standards and self-service patterns that help teams adopt it consistently. This is not a general application monitoring role; ownership of dashboards, alerts, or telemetry for every application team is not what we are looking for here.

What You will be Doing:

Platform operations and reliability: Monitor platform health, investigate and diagnose issues, and support service recovery across production and non-production environments. Contribute to service readiness through dashboards, SLOs, SLIs, runbooks, resilience checks, and performance validation. Use metrics, logs, traces, and service health data to understand issues and guide practical decisions. Incident response and continuous improvement: Support incident response, triage, problem management, and root cause analysis. Drive follow-up actions that reduce repeat issues and improve operational consistency. Improve automation and GitOps-based workflows to reduce manual work across the team. Enablement and documentation: Maintain documentation, onboarding guides, and self-service tooling so teams can operate with less reliance on direct support. Provide direct enablement on telemetry standards, alerting guidance, SLO practices, and platform workflows.

What You will Bring:

Experience supporting production platforms or services in an SRE, platform engineering, infrastructure, DevOps, or operations engineering role. Experience using observability data such as metrics, logs, traces, alerts, dashboards, or service health views to investigate issues. Understanding of incident response, problem management, service readiness, or operational support processes. Experience using automation or infrastructure-as-code practices to support repeatable delivery and operations. Solid understanding of cloud, container, Linux, networking, or distributed system environments. Ability to communicate technical information clearly to engineering, operations, and service stakeholders. Experience writing or maintaining operational documentation such as runbooks, support guides, or onboarding material. A practical approach to improving reliability, reducing manual work, and helping teams use shared platforms optimally.

Desirable skills and experience:

Experience with observability, monitoring, telemetry, or data pipeline technologies such as OpenTelemetry, Grafana, ClickHouse, Cribl, Datadog, BigPanda, Redis, Flink, or similar tools. Experience with GitOps workflows and tools such as Git, CI/CD pipelines, pull requests, environment promotion, or configuration-as-code. Experience building or supporting internal platforms used by multiple engineering teams. Experience defining or using SLOs, SLIs, error budgets, alert quality measures, or service health models. Experience supporting telemetry pipelines, data routing, data filtering, retention, or cost management. Experience working in a regulated, financial services, or large enterprise technology environment. Experience helping engineering teams adopt shared standards, templates, or self-service platform capabilities.

What You'll Get in return:

This is a phenomenal opportunity to work on a platform that directly improves how teams across LSEG understand and operate their systems!

Career Stage: Senior Associate London Stock Exchange Group (LSEG)

Information: Join us and be part of a team that values innovation, quality, and continuous improvement. If you're ready to take your career to the next level and make a significant impact, we'd love to hear from you.

LSEG is a leading global financial markets infrastructure and data provider. Our purpose is driving financial stability, empowering economies and enabling customers to create sustainable growth. Our purpose is the foundation on which our culture is built. Our values of Integrity, Partnership, Excellence and Change underpin our purpose and set the standard for everything we do, every day. They go to the heart of who we are and guide our decision making and everyday actions.

We are proud to be an equal opportunities employer. This means that we do not discriminate on the basis of anyone's race, religion, colour, national origin, gender, sexual orientation, gender identity, gender expression, age, marital status, veteran status, pregnancy or disability, or any other basis protected under applicable law. Conforming with applicable law, we can reasonably accommodate applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs.

You will be part of a collaborative and creative culture where we encourage new ideas. We are committed to sustainability across our global business and we are proud to partner with our customers to help them meet their sustainability objectives.

Our charity, the LSEG Foundation provides charitable grants to community groups that help people access economic opportunities and build a secure future with financial independence. Colleagues can get involved through fundraising and volunteering.

LSEG offers a range of tailored benefits and support, including healthcare, retirement planning, paid volunteering days and wellbeing initiatives.

  • healthcare
  • retirement planning
  • paid volunteering days
  • wellbeing initiatives
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

DevOps Engineer
DevOps Engineer

London Stock Exchange Group • Creve Coeur (MO)

On-site
USD 95,000 - 150,000
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

London Stock Exchange Group • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

London Stock Exchange Group • St. Louis (MO)

On-site
USD 140,000 - 180,000
Healthcare
Retirement planning
Paid volunteering days
+1
Senior Platform Engineer
Senior Platform Engineer

LSEG • St. Louis (MO)

On-site
USD 120,000 - 160,000
Senior Platform Engineer
Senior Platform Engineer

London Stock Exchange Group • St. Louis (MO)

On-site
USD 120,000 - 180,000
Healthcare
Retirement planning
Paid volunteering days
Site Reliability Engineer - - Electronic Trading Team
Site Reliability Engineer - - Electronic Trading Team

London Stock Exchange Group • St. Louis (MO)

On-site
USD 120,000 - 180,000
Healthcare
Retirement planning
Paid volunteering days
+1
Senior Engineer, Site Reliability Engineer
Senior Engineer, Site Reliability Engineer

London Stock Exchange Group • Creve Coeur (MO)

On-site
USD 140,000 - 200,000
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG • Raleigh (NC)

On-site
USD 140,000 - 190,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

London Stock Exchange Group • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Senior Manager, Service Engineering
Senior Manager, Service Engineering

London Stock Exchange Group • New York (NY)

On-site
USD 114,000 - 190,000