Lead Observability Engineer

Tria

City Of London

On-site

GBP 111,000 - 148,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Tria in London is seeking a Senior Observability Engineer / SRE Lead to assess and define observability roadmaps across a complex hybrid estate. You will map telemetry sources, monitoring tools, dashboards and ownership across services, platforms, infrastructure and networks.

In this contract role, you will deliver meaningful service health views, identify telemetry gaps and implement pragmatic improvements.

Qualifications

  • Proven experience leading enterprise-scale observability initiatives.
  • Hands-on Grafana, Grafana Cloud, OpenTelemetry and modern telemetry pipelines.
  • Deep understanding of metrics, logs, traces, SLIs, SLOs and error budgets.
  • Experience designing observability across cloud and on-prem platforms.
  • Strong knowledge of Azure observability tooling (Monitor, Log Analytics, Application Insights).

Responsibilities

  • Discover and document existing telemetry sources, monitoring tools, dashboards and ownership.
  • Work with technical teams and suppliers to gain access to telemetry.
  • Deliver meaningful dashboards and service health views.
  • Identify gaps in telemetry, monitoring and alerting and implement improvements.
  • Define health indicators for critical business journeys.
  • Introduce modern observability practices (SLIs, SLOs, burn-rate alerting).
  • Develop a roadmap for OpenTelemetry adoption.
  • Assess options for a centralised telemetry platform (Grafana Cloud).
  • Evaluate tooling rationalisation, operating costs and telemetry economics.
  • Define an observability target architecture, standards and implementation roadmap.

Skills

Grafana
OpenTelemetry
Telemetry pipelines
SLIs & SLOs
Error budgets
Cloud & on-prem
Leadership

Tools

Grafana Cloud
Azure Monitor
Log Analytics
Application Insights
Distributed tracing

Job description

Location: London, onsite 3 days per week (Sheffield as an alternative)

Rate: £tbd/day inside IR35

Duration: 6 months+

Are you a Senior Observability Engineer / SRE Lead, with demonstrable experience of assessing and defining observability and monitoring roadmaps within enterprise scale environments?

The Lead Observability Engineer / SRE Lead will be required to assess a complex hybrid estate, understand how services, platforms, infrastructure and networks should be monitored, and work across multiple internal teams, partners and suppliers to build a consolidated view of existing telemetry, monitoring and alerting capabilities.

As well as short term tactical objectives, the role will also be focussed on longer term strategic ones.

Responsibilities
  • Discover and document existing telemetry sources, monitoring tools, dashboards and ownership
  • Work with technical teams and suppliers to gain access to telemetry
  • Deliver meaningful dashboards and service health views
  • Identify gaps in telemetry, monitoring and alerting coverage and implement pragmatic improvements
  • Define health indicators for critical business journeys
  • Introduce modern observability practices where practical, including SLIs, SLOs and a roadmap towards burn-rate alerting
  • Develop a roadmap for OpenTelemetry adoption
  • Assess options for a centralised telemetry platform, including Grafana Cloud
  • Evaluate tooling rationalisation opportunities, operating costs and telemetry economics
  • Define an observability target architecture, standards and implementation roadmap
Requirements
  • Proven experience leading enterprise-scale observability initiatives
  • Strong hands‑on expertise with Grafana, Grafana Cloud, OpenTelemetry and modern telemetry pipelines
  • Deep understanding of metrics, logs, traces, distributed tracing, alerting, SLIs, SLOs and error‑budget concepts
  • Experience designing observability solutions across cloud PaaS, IaaS, on‑premises, legacy and third‑party hosted platforms
  • Strong knowledge of Azure observability tooling, including Azure Monitor, Log Analytics and Application Insights

Lead Observability Engineer / SRE Lead / Lead Site Reliability Engineer

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Lead Observability Engineer
Lead Observability Engineer

Tria • Greater London

On-site
GBP 92,000 - 148,000
Lead Observability Engineer
Lead Observability Engineer

TRIA Recruitment Limited • Greater London

On-site
GBP 90,000 - 120,000
Senior Observability Engineer & SRE Lead
Senior Observability Engineer & SRE Lead

TRIA Recruitment Limited • Greater London

On-site
GBP 90,000 - 120,000
Observability SRE
Observability SRE

HCLTech • Greater London

On-site
GBP 70,000 - 95,000
Senior Platform & Observability Engineer
Senior Platform & Observability Engineer

Koda Tech • Greater London

Hybrid
GBP 90,000 - 120,000
Senior Site Reliability Engineer (LON)
Senior Site Reliability Engineer (LON)

McNally Recruitment Ltd • Greater London

Hybrid
GBP 90,000 - 150,000
Benefits as Cash
Hybrid work model
Site Reliability Engineer
Site Reliability Engineer

The Business Connection Group • Greater London

Remote
GBP 130,000 - 150,000
Enterprise Observability Lead & SRE Roadmap Architect
Enterprise Observability Lead & SRE Roadmap Architect

Tria • City Of London

On-site
GBP 111,000 - 148,000
Senior SRE: OpenTelemetry & Observability (Hybrid London)
Senior SRE: OpenTelemetry & Observability (Hybrid London)

EPAM Systems, Inc. • Greater London

Hybrid
GBP 100,000 - 140,000
ESPP
Private medical insurance
Dental care
+4
Lead Observability Engineer | Enterprise SRE Leader
Lead Observability Engineer | Enterprise SRE Leader

Tria • Greater London

On-site
GBP 92,000 - 148,000