Site Reliability Engineer III

LexisNexis Risk Solutions (Europe) Limited Company

Dublin

On-site

EUR 51,000 - 86,000

Full time

28 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Country-specific benefits

Job summary

LexisNexis Risk Solutions is seeking a Site Reliability Engineer to design, operate and continuously improve secure Azure platforms that support large-scale production environments.

You will provide technical leadership for reliability initiatives, develop reusable automation and partner with engineering, security and product teams to enhance resilience and operational performance. Strong Azure, Kubernetes and IaC skills are required.

Qualifications

  • Experience in Site Reliability Engineering or Cloud/Platform Engineering for large-scale production environments.
  • Expert knowledge of Microsoft Azure and Kubernetes (AKS preferred).
  • Advanced Terraform skills and reusable IaC design.
  • Strong Git/GitOps with GitHub Actions and Azure DevOps.
  • Scripting in Python/PowerShell/Go/Bash; Linux administration.
  • Experience with monitoring, observability and distributed tracing.
  • Incident management, RCA & capacity planning; disaster recovery.
  • Leadership, problem solving, and strong communication; automation-first mindset.

Responsibilities

  • Design, operate and improve guaranteed Azure-based platforms for reliability.
  • Define and maintain SLIs/SLOs and error budgets to improve resilience.
  • Lead RCAs and post-incident reviews; implement auto-scaling and multi-region failover.
  • Build IaC with Terraform; automate operational tasks with Python/PowerShell/Go/Bash.
  • Manage AKS clusters, container security, and GitOps with ArgoCD.
  • Develop monitoring, logging, alerts and tracing; dashboards in Grafana/Azure Monitor.
  • Create CI/CD pipelines with GitHub Actions/Azure DevOps; blue/green canaries.
  • Provide technical leadership and mentoring for SRE engineers; participate in on-call.

Skills

SRE experience
Cloud engineering
DevOps
Incident management
Technical leadership
Automation-first approach
Communication & collaboration

Tools

Terraform
GitHub Actions
Azure DevOps
ArgoCD
Grafana

Job description

About the Business

Are you passionate about building reliable, secure and scalable cloud platforms? Would you like to lead Azure reliability initiatives and help engineering teams adopt Site Reliability Engineering best practices? LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Insurance vertical, we provide customers with solutions and decision tools that combine public and industry specific content with advanced technology and analytics to assist them in evaluating and predicting risk and enhancing operational efficiency. Our insurance risk solutions help drive better data-driven decisions across the insurance policy lifecycle, all while reducing risk. You can learn more about LexisNexis Risk at https://risk.lexisnexis.com/insurance.

About the Role

As a Site Reliability Engineer, you will design, operate and continuously improve secure, highly available Azure platforms that support large-scale production environments. You will provide technical leadership for reliability initiatives, develop reusable automation and partner with engineering, security, architecture and product teams to improve resilience and operational performance.

Responsibilities
  • Design, implement and maintain highly available Azure infrastructure, including Azure Kubernetes Service, Virtual Machines, Functions, App Services, Storage, Networking, Key Vault and Azure Monitor.
  • Define and maintain Service Level Indicators, Service Level Objectives and Error Budgets, driving continuous improvements in platform reliability and availability.
  • Lead root cause analysis and post-incident reviews, developing resilience patterns such as auto-scaling, self-healing, disaster recovery, multi-region failover and high-availability architectures.
  • Design and maintain Infrastructure as Code using Terraform, build reusable platform components and automate manual operational processes using Python, PowerShell, Go or Bash.
  • Operate and optimise AKS clusters, establish container security standards, implement GitOps practices using tools such as ArgoCD and manage upgrades, capacity and workloads.
  • Develop monitoring, logging, alerting and distributed-tracing strategies, maintaining dashboards through Grafana and Azure Monitor to identify service degradation proactively.
  • Build and support reliable, repeatable and auditable CI/CD pipelines using GitHub Actions and Azure DevOps, including blue/green and canary deployment strategies.
  • Provide technical leadership, mentor SRE I and SRE II engineers, promote SRE practices across teams and participate in on-call rotations and major incident management.
Requirements
  • Experience in Site Reliability Engineering, Cloud Engineering, DevOps or Platform Engineering, including supporting large-scale production environments.
  • Expert knowledge of Microsoft Azure and strong experience with Kubernetes, preferably Azure Kubernetes Service.
  • Advanced Terraform skills and experience designing reusable Infrastructure as Code components.
  • Strong Git and GitOps practices, with experience using CI/CD platforms such as GitHub Actions and Azure DevOps.
  • Strong Linux administration, scripting and programming skills using languages such as Python, PowerShell, Go or Bash.
  • Experience with monitoring and observability platforms, distributed tracing, application performance monitoring and proactive alerting.
  • Skills in incident management, root cause analysis, capacity planning, performance optimisation, disaster recovery testing and production operations.
  • Technical leadership, strategic thinking, problem solving, effective communication, collaboration, customer focus and an automation-first approach.
  • Azure, Kubernetes or Terraform certifications are welcomed.

Learn more about the LexisNexis Risk team and how we work here.

Primary Location

Base Pay Range: Ireland - Dublin (Rockfield Central) €51,400 - €85,700.

We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits.

Click here to access benefits specific to your location.

We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact 1-855-833-5120.

We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.

USA Job Seekers: EEO Know Your Rights.

RELX is a global provider of information-based analytics and decision tools for professional and business customers, enabling them to make better decisions, get better results and be more productive. Our purpose is to benefit society by developing products that help researchers advance scientific knowledge; doctors and nurses improve the lives of patients; lawyers promote the rule of law and achieve justice and fair results for their clients; businesses and governments prevent fraud; consumers access financial services and get fair prices on insurance; and customers learn about markets and complete transactions. Our purpose guides our actions beyond the products that we develop. It defines us as a company. Every day across RELX our employees are inspired to undertake initiatives that make unique contributions to society and the communities in which we operate.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

RELX INC • Ireland

Remote
EUR 49,000 - 82,000
Site Reliability Engineer
Site Reliability Engineer

Talent Octopusventures • Ireland

Remote
EUR 49,000 - 82,000
Site Reliability Engineer
Site Reliability Engineer

LexisNexis Risk Solutions • Ireland

Remote
EUR 49,000 - 82,000
Principal Data Scientist
Principal Data Scientist

LexisNexis Risk Solutions (Europe) Limited Company • Dublin

On-site
EUR 75,000 - 126,000
Senior Software Engineer I
Senior Software Engineer I

Talent Octopusventures • Dublin

On-site
EUR 57,000 - 94,000
Senior Software Engineer I
Senior Software Engineer I

RELX • Dublin

On-site
EUR 57,000 - 94,000
Azure SRE Lead: Reliability, Automation & CI/CD
Azure SRE Lead: Reliability, Automation & CI/CD

LexisNexis Risk Solutions (Europe) Limited Company • Dublin

On-site
EUR 51,000 - 86,000
Country-specific benefits
Cloud Data Platform Engineer
Cloud Data Platform Engineer

RELX • Dublin

On-site
EUR 47,000 - 78,000
Principal Data Scientist
Principal Data Scientist

RELX INC • Dublin

On-site
EUR 75,000 - 126,000
Senior Cloud Database Engineer II
Senior Cloud Database Engineer II

LexisNexis Risk Solutions (Europe) Limited Company • Ireland

Remote
EUR 68,000 - 114,000