Technical Lead - Site Reliability Engineering

London Stock Exchange Group

Town of Texas (WI)

On-site

USD 140,000 - 190,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

London Stock Exchange Group (LSEG) is evolving its Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across its Markets and Risk Intelligence division. As a Technical Lead SRE, you will be a senior hands-on technical leader shaping reliability foundations for both new and existing platforms.

You will collaborate across Architecture, Engineering, Security and Platform teams, mentor engineers, and drive incident readiness, cost

Qualifications

  • Bachelor’s Degree in Computer Science or related field.
  • 10+ years of hands-on technical experience in SRE, Platform Engineering, Infrastructure, or related roles.
  • Strong experience with Azure, including AKS, Azure Container Apps, Virtual Machines, VNet, and Azure managed services.
  • Hands-on experience with Kubernetes and containerised platforms.
  • Strong background in Linux systems administrations.
  • Proven experience designing and operating observability platforms, including monitoring, logging, and alerting.
  • Hands-on experience with Datadog for metrics, logs, APM, and alerting.
  • Strong understanding of SRE principles, including SLOs, error budgets, incident management, and reliability engineering.
  • Experience working closely with architecture and engineering teams on system design and delivery.
  • Solid understanding of cloud security principles and experience collaborating with security teams.
  • Experience with cloud cost optimisation strategies and tooling.
  • Experience integrating AI with observability stacks (Prometheus, Grafana, ELK, OpenTelemetry) for proactive issue detection.

Responsibilities

  • Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.
  • Collaborate with Architecture and Engineering teams to embed reliability, scalability, security, and observability into system design.
  • Define, implement, and champion observability standards, tooling, and guidelines across metrics, logs, traces, and SLIs/SLOs.
  • Design and evolve monitoring and alerting solutions to improve visibility and reduce toil.
  • Drive reliability improvements through incident reduction, performance tuning, and resilient patterns.
  • Partner with Security teams to meet compliance, security, and risk-management expectations.
  • Lead handovers from project delivery into BAU SRE operations with documentation and readiness.
  • Influence design decisions through data-driven cloud cost optimization and efficiency initiatives.
  • Mentor engineers and foster a culture of learning and development.

Skills

SRE
Platform Engineering
Azure
Kubernetes
Linux
Observability
Datadog
SLOs
Incident Management
IaC
Cloud Security
AI in Observability
AWS
OpenTelemetry
Prometheus
Grafana

Education

Bachelor’s Degree in Computer Science

Tools

Datadog
Kubernetes
Prometheus
Grafana
ELK
OpenTelemetry
Terraform
CloudFormation
AWS

Job description

Role Profile

We are evolving our Site Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division. As a Technical Lead SRE, you will be a senior hands-on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate with Architecture, Engineering, Security, and Platform teams to ensure reliability is built into systems from day one. You will work closely with global teams and may occasionally be called upon for major incidents or critical issues. This position requires a highly proactive, hard-working expert with strong leadership presence and ownership of platform reliability outcomes.

Key Responsibilities

We are looking for a person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do! Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one. Collaborate with Architecture and Engineering teams to embed reliability, scalability, security, and observability into system design. Define, implement, and champion observability standards, tooling, and guidelines across metrics, logs, traces, and SLIs/SLOs. Design and evolve monitoring and alerting solutions that improve visibility, reduce toil, and strengthen system health. Continuously drive reliability improvements across our environments through incident reduction, performance tuning, and building resilient patterns. Partner with Security teams to ensure our platforms meet compliance, security, and risk-management expectations. Lead seamless handovers from project delivery into BAU SRE operations by ensuring documentation, readiness, and strong operational practices. Influence architectural and design decisions through data-driven cloud cost optimization and efficiency initiatives. Be a technical leader and mentor supporting engineers, shaping engineering standards, and fostering a culture of learning and development.

PERSON SPECIFICATION

Education: Bachelor’s Degree in Computer Science or related field

Required skills and experience 10+ years of hands-on technical experience in SRE, Platform Engineering, Infrastructure, or related roles

Strong experience with Azure, including services such as AKS, Azure Container Apps, Virtual Machines, virtual networking (VNet), and Azure managed services

Hands-on experience with Kubernetes and containerised platforms

Strong background in Linux systems administrations.

Proven experience designing and operating observability platforms, including monitoring, logging, and alerting

Hands-on experience with Datadog for metrics, logs, APM, and alerting

Strong understanding of SRE principles, including SLOs, error budgets, incident management, and reliability engineering

Experience working closely with architecture and engineering teams on system design and delivery

Solid understanding of cloud security principles and experience collaborating with security teams

Experience with cloud cost optimisation strategies and tooling

Experience integrating AI with observability stacks (Prometheus, Grafana, ELK, OpenTelemetry) for proactive issue detection.

Good to have Skills Experience or working knowledge of AWS Experience supporting multi-cloud or hybrid environments Exposure to Infrastructure as Code (e.g., Terraform, CloudFormation) Experience in large-scale, complex, or regulated environments Knowledge of vector databases and RAG architectures for building internal SRE knowledge assistants. Knowledge of Generative AI and LLM platforms (e.g., Claude, Amazon Bedrock)

Person Specification Strong technical authority with the ability to influence design and operational decisions Highly collaborative, comfortable working across architecture, engineering, security, and operations teams Calm and methodical under pressure, especially during incidents and critical issues Pragmatic problem-solver who balances reliability, security, cost, and delivery speed Clear communicator, able to explain complex technical concepts to diverse audiences.

Career Stage: Senior Associate, London Stock Exchange Group (LSEG) Information

Join us and be part of a team that values innovation, quality, and continuous improvement. If you're ready to take your career to the next level and make a significant impact, we'd love to hear from you.

LSEG is a leading global financial markets infrastructure and data provider. Our purpose is driving financial stability, empowering economies and enabling customers to create sustainable growth.

Our purpose is the foundation on which our culture is built. Our values of Integrity, Partnership, Excellence and Change underpin our purpose and set the standard for everything we do, every day. They go to the heart of who we are and guide our decision making and everyday actions.

Working with us means that you will be part of a dynamic organisation of 25,000 people across 65 countries. However, we will value your individuality and enable you to bring your true self to work so you can help enrich our diverse workforce.

We are proud to be an equal opportunities employer. This means that we do not discriminate on the basis of anyone’s race, religion, colour, national origin, gender, sexual orientation, gender identity, gender expression, age, marital status, veteran status, pregnancy or disability, or any other basis protected under applicable law. Conforming with applicable law, we can reasonably accommodate applicants' and employees' religious practices and beliefs, as well as mental health or physical disability needs.

You will be part of a collaborative and creative culture where we encourage new ideas. We are committed to sustainability across our global business and we are proud to partner with our customers to help them meet their sustainability objectives.

Our charity, the LSEG Foundation provides charitable grants to community groups that help people access economic opportunities and build a secure future with financial independence. Colleagues can get involved through fundraising and volunteering.

LSEG offers a range of tailored benefits and support, including healthcare, retirement planning, paid volunteering days and wellbeing initiatives.

LSEG (London Stock Exchange Group) is a leading global financial markets infrastructure and data provider. Our purpose is driving financial stability, empowering economies and enabling customers to create sustainable growth. Our culture of connecting, creating opportunity and delivering excellence shapes how we think, how we do things and how we help our people fulfil their potential.

Our Data & Analytics, Capital Markets and Post Trade divisions have a combined power that provides a comprehensive, integrated suite of trusted financial market infrastructure services to help our customers pursue their ambitions.

Explore our divisions LSEG is headquartered in the United Kingdom, with significant operations in 70 countries across Europe, the Middle East, Africa, North America, Latin America and Asia Pacific. Find out more Get to know some of our people who are pushing the boundaries of technology, finance and more around the world.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG • Raleigh (NC)

On-site
USD 140,000 - 190,000
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG (London Stock Exchange Group) • Allen (TX)

On-site
USD 170,000 - 210,000
Healthcare
Retirement planning
Volunteer days
+1
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

London Stock Exchange Group • Town of Texas (WI)

On-site
USD 140,000 - 190,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG (London Stock Exchange Group) • Raleigh (NC)

On-site
USD 120,000 - 180,000
Senior Engineer, Site Reliability Engineer
Senior Engineer, Site Reliability Engineer

London Stock Exchange Group • Creve Coeur (MO)

On-site
USD 140,000 - 200,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

London Stock Exchange Group • United States

On-site
USD 120,000 - 150,000
Healthcare
Retirement planning
Paid volunteering days
+1
Site Reliability Engineer
Site Reliability Engineer

London Stock Exchange Group • St. Louis (MO)

On-site
USD 140,000 - 180,000
Healthcare
Retirement planning
Paid volunteering days
+1
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

LSEG (London Stock Exchange Group) • Raleigh (NC)

On-site
USD 130,000 - 180,000
Technical Lead - Site Reliability Engineering
Technical Lead - Site Reliability Engineering

London Stock Exchange Group • Raleigh (NC)

On-site
USD 140,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

London Stock Exchange Group • Creve Coeur (MO)

On-site
USD 120,000 - 160,000
Healthcare
Retirement planning
Paid volunteering days
+1