SRE - Driving Reliability & Observability

RWS

Dublin

On-site

EUR 90,000 - 120,000

Full time

19 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

RWS is hiring a Site Reliability Engineer (SRE) to operate within frontier technology teams, building at pace, establishing engineering standards, and shaping architecture to improve reliability and observability across the estate.

You will work across product engineering, infrastructure operations, platform engineering, enterprise tech, and data teams, driving reliability through expertise, clarity of thinking, and hands-on delivery.

Qualifications

  • Hands-on SRE/operational engineering experience across distributed systems.
  • Expertise in observability (metrics, logging, tracing) and platforms such as Prometheus, Grafana, ELK, Datadog, Splunk, Honeycomb, or equivalents.
  • Experience with OpenTelemetry and modern telemetry pipelines.
  • Knowledge of AWS and/or GCP, Kubernetes/EKS, Linux systems, and CI/CD tooling.
  • Ability to analyse complex system behaviour, diagnose issues, and design scalable, pragmatic solutions.
  • Strong technical communication skills, able to influence through clarity, evidence, and thoughtful design.

Responsibilities

  • Act as a trusted technical partner to engineering teams, helping design and operate more resilient systems.
  • Define and drive adoption of SLIs, SLOs, and error budgets; establish reliability baselines and standards.
  • Provide SRE guidance in incident reviews, deep-dives, and long-term remediation with root-cause analysis.
  • Shape the architecture of a new observability platform through hands-on design and input.
  • Define instrumentation standards (metrics, logs, traces, events) using OpenTelemetry.
  • Collaborate on tool consolidation, scalable telemetry pipelines, and improved signal quality.
  • Build reusable frameworks and components to raise visibility and operational excellence.
  • Identify reliability bottlenecks and design automation, re-architecture, pipelines and guardrails.
  • Improve deployment stability and service quality across product lines.
  • Foster cross-functional collaboration with Product, Security, Data, and Enterprise Tech.

Skills

SRE/operational engineering
Observability expertise
Technical communication
Cross-functional collaboration

Tools

Prometheus
Grafana
ELK
Datadog
Splunk
Honeycomb
OpenTelemetry
Kubernetes
Linux
CI/CD tooling
AWS
GCP
EKS

Job description

RWS is hiring a Site Reliability Engineer (SRE) to operate within frontier technology teams, building at pace, establishing engineering standards, and shaping architecture to improve reliability and observability across the estate.

You will work across product engineering, infrastructure operations, platform engineering, enterprise tech, and data teams, driving reliability through expertise, clarity of thinking, and hands-on delivery.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

RWS • Dublin

On-site
EUR 90,000 - 120,000
Site Reliability Engineer: Build Resilient, Observability‑Driven Systems
Site Reliability Engineer: Build Resilient, Observability‑Driven Systems

RWS • Dublin

On-site
EUR 90,000 - 130,000
Site Reliability Engineer at RWS
Site Reliability Engineer at RWS

RWS • Dublin

On-site
EUR 90,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Harvey Nash • Dublin

On-site
EUR 90,000 - 130,000
Remote SRE - Automate, Scale and Observe
Remote SRE - Automate, Scale and Observe

Social Discovery Group • Ireland

Remote
EUR 90,000 - 120,000
Remote work
Vacation 28 days
Wellness days (7)
+6
SRE (Application Support + Dev-Ops + Automation)
SRE (Application Support + Dev-Ops + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
SRE Engineering Manager — Uptime & Automation Leader
SRE Engineering Manager — Uptime & Automation Leader

Google Inc. • Leinster

On-site
EUR 150,000 - 153,000
Equity
Benefits
Bonus target
SRE for AI Cloud Reliability & Automation
SRE for AI Cloud Reliability & Automation

Crusoe • Dublin

On-site
EUR 65,000 - 90,000
Pension contributions
Private health insurance
Dental insurance
+2
SRE Tech Lead: Reliability & AI-Driven Ops
SRE Tech Lead: Reliability & AI-Driven Ops

AMCS Group • Limerick

On-site
EUR 110,000 - 140,000
Security SRE Leader: Global Reliability & Automation
Security SRE Leader: Global Reliability & Automation

Lex • Dublin

On-site
EUR 150,000 - 210,000