Senior Observability Engineer: Cloud & Platform Reliability

Tata Consultancy Services

Greater London

Hybrid

GBP 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tata Consultancy Services UK&I is seeking a Senior Observability Engineer to design, implement, and support enterprise-scale observability solutions across apps, infra and cloud platforms. You will drive monitoring strategy, enable proactive incident detection, and improve platform reliability.

You will work with OpenTelemetry, Grafana Enterprise, Prometheus and OpenShift/Kubernetes, creating dashboards, alerts and scalable monitoring across a hybrid London-based environment.

Qualifications

  • Proven experience designing and operating enterprise observability solutions.
  • Hands-on with OpenTelemetry, Grafana Enterprise Stack or ITRS Geneos administration.
  • Goes with Google Cloud Observability.
  • Strong experience with OpenShift and/or Kubernetes environments.
  • Expertise in Prometheus monitoring and PromQL query development.
  • Grafana dashboard creation, alerting and data source management.
  • Experience deploying and maintaining Helm charts.
  • Understanding observability principles: metrics, logs, traces, alerting.

Responsibilities

  • Design, build and manage enterprise observability solutions covering metrics, logs and distributed tracing.
  • Support migration from legacy monitoring platforms to a modern observability ecosystem.
  • Administer, maintain and scale observability platforms on OpenShift/Kubernetes.
  • Implement and manage observability instrumentation using OpenTelemetry.
  • Develop and maintain Grafana dashboards, alerts and data sources.
  • Build and optimize Prometheus-based monitoring and craft PromQL queries.
  • Create and maintain Helm charts for deployment and lifecycle management.
  • Automate deployments and configuration through scripting and tooling.
  • Collaborate with engineering and infra teams to set standards and best practices.
  • Contribute to platform reliability and continuous service improvement.

Skills

Observability design
OpenTelemetry
Grafana Enterprise
Prometheus/PromQL
OpenShift administration
Kubernetes administration
Helm charts
Python scripting
Bash scripting
Cloud observability
SRE concepts

Tools

OpenShift
Kubernetes
Grafana
Prometheus
Loki
Tempo
Mimir
OpenTelemetry
Helm

Job description

Tata Consultancy Services UK&I is seeking a Senior Observability Engineer to design, implement, and support enterprise-scale observability solutions across apps, infra and cloud platforms. You will drive monitoring strategy, enable proactive incident detection, and improve platform reliability.

You will work with OpenTelemetry, Grafana Enterprise, Prometheus and OpenShift/Kubernetes, creating dashboards, alerts and scalable monitoring across a hybrid London-based environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Monitoring Engineer
Monitoring Engineer

Tata Consultancy Services • Greater London

Hybrid
GBP 90,000 - 120,000
Observability Engineer - Cloud Telemetry & Automation
Observability Engineer - Cloud Telemetry & Automation

develop • Greater London

On-site
Observability SME
Observability SME

Gazelle Global Consulting Limited • Greater London

On-site
GBP 148,000 - 221,000
Excellent day rate
Observability Engineer (Dynatrace) — Telemetry & Performance
Observability Engineer (Dynatrace) — Telemetry & Performance

Computacenter AG & Co. oHG • Greater London

On-site
GBP 75,000 - 110,000
Observability Engineer: Dynatrace, Grafana & Cloud
Observability Engineer: Dynatrace, Grafana & Cloud

Computacenter • Greater London

On-site
GBP 50,000 - 70,000
Observability Platform Engineer — Hybrid London
Observability Platform Engineer — Hybrid London

N Consulting Limited • Greater London

Hybrid
GBP 60,000 - 80,000
Senior Observability Engineer: Honeycomb & OpenTelemetry
Senior Observability Engineer: Honeycomb & OpenTelemetry

IG Group • Greater London

On-site
GBP 90,000 - 130,000
Competitive salary
Private medical cover for you and your
Life insurance
+2
Senior Observability Engineer – OpenShift, Grafana, GCP
Senior Observability Engineer – OpenShift, Grafana, GCP

LTM • Belfast

On-site
GBP 70,000 - 100,000
Senior Observability Engineer — Scale Reliability & Telemetry
Senior Observability Engineer — Scale Reliability & Telemetry

IG Group • City Of London

On-site
GBP 110,000 - 150,000
Competitive salary
Private medical cover
Flexible benefits package
+1
Senior SRE: Cloud Reliability & Observability Leader
Senior SRE: Cloud Reliability & Observability Leader

Omilia • United Kingdom

On-site
GBP 90,000 - 130,000
Fixed compensation
Long-term employment
Professional growth
+1