Site Reliability Engineer

RWS

Dublin

On-site

EUR 90,000 - 120,000

Full time

8 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

RWS is hiring a Site Reliability Engineer (SRE) to operate within frontier technology teams, building at pace, establishing engineering standards, and shaping architecture to improve reliability and observability across the estate.

You will work across product engineering, infrastructure operations, platform engineering, enterprise tech, and data teams, driving reliability through expertise, clarity of thinking, and hands-on delivery.

Qualifications

  • Hands-on SRE/operational engineering experience across distributed systems.
  • Expertise in observability (metrics, logging, tracing) and platforms such as Prometheus, Grafana, ELK, Datadog, Splunk, Honeycomb, or equivalents.
  • Experience with OpenTelemetry and modern telemetry pipelines.
  • Knowledge of AWS and/or GCP, Kubernetes/EKS, Linux systems, and CI/CD tooling.
  • Ability to analyse complex system behaviour, diagnose issues, and design scalable, pragmatic solutions.
  • Strong technical communication skills, able to influence through clarity, evidence, and thoughtful design.

Responsibilities

  • Act as a trusted technical partner to engineering teams, helping design and operate more resilient systems.
  • Define and drive adoption of SLIs, SLOs, and error budgets; establish reliability baselines and standards.
  • Provide SRE guidance in incident reviews, deep-dives, and long-term remediation with root-cause analysis.
  • Shape the architecture of a new observability platform through hands-on design and input.
  • Define instrumentation standards (metrics, logs, traces, events) using OpenTelemetry.
  • Collaborate on tool consolidation, scalable telemetry pipelines, and improved signal quality.
  • Build reusable frameworks and components to raise visibility and operational excellence.
  • Identify reliability bottlenecks and design automation, re-architecture, pipelines and guardrails.
  • Improve deployment stability and service quality across product lines.
  • Foster cross-functional collaboration with Product, Security, Data, and Enterprise Tech.

Skills

SRE/operational engineering
Observability expertise
Technical communication
Cross-functional collaboration

Tools

Prometheus
Grafana
ELK
Datadog
Splunk
Honeycomb
OpenTelemetry
Kubernetes
Linux
CI/CD tooling
AWS
GCP
EKS

Job description

We are hiring a Site Reliability Engineer (SRE) to operate within RWS’s frontier technology teams, building at pace, establishing engineering standards, and helping shape the architecture and practises needed to improve reliability and observability across our entire estate.

This role is a technical position, focused on building reliability into fast-paced product development. You will work across product engineering, infrastructure operations, platform engineering, enterprise tech, and data teams, driving reliability through expertise, influence, clarity of thinking, and hands‑on delivery.

About Product & Technology

Product & Technology plays a pivotal role in aligning the organization with its strategic objectives and enhancing shareholder value. Product & Technology is responsible for establishing unified standards and governance practices throughout the company. Additionally, we oversee the development and maintenance of core applications essential for the seamless operation of various functions across the organization. We are committed to driving and executing future roadmaps that are in line with the overall strategic direction of RWS.

With a global reach, Product & Technology provides support services to over 7500 end users worldwide. We take pride in managing the information security operation and safeguarding all our assets. Our core functions encompass Enterprise & Technical Architecture, Network & Voice, Infrastructure, Service Delivery, Service Operations, Data & Analytics, Security & Quality Compliance, Transformation, Application Development, Enterprise Platforms. With a dedicated team of over 500 staff, Product & Technology ensures a strong presence across all regions, enabling efficient and effective support to our global operations.

Key Responsibilities
Technical Reliability
  • Act as a trusted technical partner to leading-edge engineering teams, helping them design and operate more resilient systems.
  • Help define and drive adoption of SLIs, SLOs, and error budgets; establish reliability baselines and engineering standards.
  • Provide SRE guidance in incident reviews, deep-dives, and long‑term remediation, ensuring real root causes are addressed.
Observability Platform Foundation
  • Help shape the architecture of RWS’s new observability platform through hands‑on design, prototyping, and technical decision input.
  • Assist with the definition of instrumentation standards (metrics, logs, traces, events) using modern, vendor‑neutral approaches such as OpenTelemetry.
  • Collaborate with platform engineering on tool consolidation, scalable telemetry pipelines, and improved signal quality.
  • Build reusable frameworks and components that improve visibility and operational excellence across teams.
Engineering Excellence & Automation
  • Identify systemic reliability bottlenecks and design technical solutions — automation, re‑architecture, pipelines, guardrails, etc.
  • Improve deployment stability, operational readiness, and the quality of services across multiple product lines.
  • Introduce resilience techniques such as load testing, chaos testing, and failure‑mode analysis where appropriate.
  • Produce clear technical documentation, runbooks, and patterns that raise engineering maturity.
Cross-Functional Technical Collaboration
  • Work closely with engineering teams across RWS’s diverse product and platform landscape to embed SRE thinking.
  • Collaborate with Product, Security, Data, and Enterprise Tech to address cross‑cutting reliability work.
  • Help define how the future Reliability & Operations organisation should operate from a technical practice perspective.
Skills & Experience

You are an experienced SRE or platform engineer with strong architectural instincts, deep operational experience, and the ability to lead complex technical improvement initiatives without formal authority.

You will have:
  • Hands‑on SRE/operational engineering experience across distributed systems.
  • Expertise in observability (metrics, logging, tracing) and platforms such as Prometheus, Grafana, ELK, Datadog, Splunk, Honeycomb, or equivalents.
  • Experience with OpenTelemetry and modern telemetry pipelines.
  • Knowledge of AWS and/or GCP, Kubernetes/EKS, Linux systems, and CI/CD tooling.
  • Ability to analyse complex system behaviour, diagnose issues, and design scalable, pragmatic solutions.
  • Strong technical communication skills, able to influence through clarity, evidence, and thoughtful design.
  • A collaborative, enabling mindset: you lift teams by helping them solve hard problems and adopt better practices.

Life at RWS - If you like the idea of working with smart people who are passionate about growing the value of ideas, data and content by making sure organizations are understood, then you’ll love life at RWS.

Our purpose is to unlock global understanding. This means our work fundamentally recognizes the value of every language and culture. So, we celebrate difference, we are inclusive and believe that diversity makes us strong. We want every employee to grow as an individual and excel in their career.

In return, we expect all our people to live by the values that unite us: to partner, putting clients fist and winning together to pioneer, innovating fearlessly and leading with vision and courage, to progress, aiming high and growing through actions and to deliver, owning the outcome and building trust with our colleagues and clients.

RWS embraces DEI

and promotes equal opportunity, we are an Equal Opportunity Employer and prohibit discrimination and harassment of any kind. RWS is committed to the principle of equal employment opportunity for all employees and to providing employees with a work environment free of discrimination and harassment. All employment decisions at RWS are based on business needs, job requirements and individual qualifications, without regard to race, religion, nationality, ethnicity, sex, age, disability, or sexual orientation. RWS will not tolerate discrimination based on any of these characteristics.

RWS Values

Get the 3Ps right – Partner, Pioneer, Progress – and we´ll Deliver together as RWS.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer at RWS
Site Reliability Engineer at RWS

RWS • Dublin

On-site
EUR 90,000 - 130,000
SRE - Driving Reliability & Observability
SRE - Driving Reliability & Observability

RWS • Dublin

On-site
EUR 90,000 - 120,000
Security Engineer
Security Engineer

RWS • Dublin

On-site
EUR 60,000 - 90,000
Senior AI Engineer
Senior AI Engineer

RWS • Dublin

On-site
EUR 110,000 - 150,000
Software Engineer Train AI (Mid-Level)
Software Engineer Train AI (Mid-Level)

RWS • Dublin

On-site
EUR 70,000 - 110,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Dublin

On-site
EUR 110,000 - 150,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Harvey Nash • Dublin

On-site
EUR 90,000 - 130,000
Site Reliability Engineering Technical Lead
Site Reliability Engineering Technical Lead

AMCS Group • Leinster

On-site
EUR 90,000 - 130,000
SRE (Application Support + Dev-Ops + Automation)
SRE (Application Support + Dev-Ops + Automation)

Fulcrum Digital • Dublin

On-site
EUR 90,000 - 120,000
Site Reliability Engineer: Build Resilient, Observability‑Driven Systems
Site Reliability Engineer: Build Resilient, Observability‑Driven Systems

RWS • Dublin

On-site
EUR 90,000 - 130,000