Site Reliability Engineering Lead

RELX INC

Philadelphia (Philadelphia County)

Remote

USD 118.000 - 220.000

Vollzeit

Vor 7 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

RELX INC in the United States is seeking a senior manager to lead the Core SRE Team within Business Services. You will provide line management to a small to medium sized group of engineers, driving reliability across cloud-native environments and mentoring engineers to grow.

The role emphasizes prioritizing tasks, enabling automation, leading post-mortem reviews, and delivering scalable, secure platforms with cost-conscious operations.

Qualifikationen

  • Expert knowledge of Kubernetes at scale.
  • Expert Terraform IaC design and multi-environment provisioning.
  • Deep knowledge of Azure Cloud services and cost optimization.
  • Experience designing and scaling CI/CD pipelines with GitHub Actions.
  • Experience with observability platforms such as Prometheus, Grafana, OpenTelemetry.
  • Strong automation skills to reduce toil with self-healing systems.
  • Advanced Python, Bash, and/or PowerShell for tooling.
  • Deep networking knowledge: TCP/IP, DNS, load balancing, VPNs, cloud networking.
  • Experience in SRE/DevOps or Infrastructure roles and leading engineering teams.
  • Proven track record leading incident response and reliability improvements.

Aufgaben

  • Manage, mentor, and grow a team of SREs; conduct 1:1s, performance reviews, and career development planning
  • Own hiring, onboarding, and team capacity/resourcing decisions
  • Set team goals, prioritize backlog, and drive planning
  • Foster a blameless post-incident culture and cross-team collaboration with Dev, Security, and Product
  • Lead reliability initiatives across infrastructure and services
  • Drive incident response activities and continuous service improvement
  • Champion automation and operational excellence across the platform
  • Support the development of scalable, secure, and resilient cloud-native environments

Kenntnisse

Kubernetes
Terraform
Azure Cloud
CI/CD with GitHub Actions
Observability (Prometheus/Grafana/Open

Tools

GitHub Actions
Prometheus
Grafana
OpenTelemetry

Jobbeschreibung

Florida

Home based-Connecticut

Home based-New Jersey

Home based-New York

Home based-Pennsylvania

Home based-Delaware

Home based-Washington DC

Home based-South Carolina

Full time

Are you passionate about building reliable, scalable platforms and helping engineering teams thrive?

Do you enjoy leading high-performing teams while driving automation, resilience, and operational excellence across modern cloud environments?

About the Business

LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Business Services vertical, we offer a multitude of solutions focused on helping businesses of all sizes drive higher revenue growth, maximize operational efficiencies, and improve customer experience. Our solutions help our customers solve difficult problems in the areas of Anti-Money Laundering/Counter Terrorist Financing, Identity Authentication & Verification, Fraud and Credit Risk mitigation and Customer Data Management. You can learn more about LexisNexis Risk at https://risk.lexisnexis.com/

About Our Team

You will be joining the Core SRE Team in Business Services, a team that oversees all the applications and infrastructure in the biggest business unit in LexisNexis Risk Solutions. We build cloud environments, migrate on prem applications to the cloud, work with self-hosted and 3rd party solutions. The successful candidate is a self-starter who assesses the situation, collaboratively develops a solution, and takes the initiative to improve performance, cost, and reliability at each opportunity.

About the Role

This is a professional management level role. Individuals are required to provide line management to a small to medium sized team of engineers, including authority over performance management, pay, and recruitment. They will ensure that tasks and projects are prioritized appropriately and provide support to team members when tasks are blocked. They will support engineers in their personal development and ensure they are working within the SRE framework. They will lead the post-mortem reviews and the timely production of RCAs. They address issues with impact beyond their own team based on knowledge of related disciplines.

Responsibilities
  • Manage, mentor, and grow a team of SREs; conduct 1:1s, performance reviews, and career development planning
  • Own hiring, onboarding, and team capacity/resourcing decisions
  • Set team goals, prioritize backlog, and drive planning
  • Foster a blameless post-incident culture and cross-team collaboration with Dev, Security, and Product
  • Lead reliability initiatives across infrastructure and services
  • Drive incident response activities and continuous service improvement
  • Champion automation and operational excellence across the platform
  • Support the development of scalable, secure, and resilient cloud-native environments
Requirements
  • Expert knowledge of Kubernetes, including cluster architecture, upgrades, autoscaling, security hardening, and troubleshooting at scale
  • Expert experience with Terraform, including modular IaC design, state management, multi-environment provisioning, and policy-as-code
  • Deep knowledge of Azure Cloud, including compute, networking, identity (AAD), storage, and cost optimization
  • Experience designing and scaling CI/CD pipelines using GitHub Actions, release strategies, and rollback automation
  • Experience with observability platforms including Prometheus, Grafana, OpenTelemetry, and SLO/SLA/error-budget management
  • Strong automation skills, focused on eliminating toil through self-healing systems and infrastructure automation
  • Advanced proficiency in Python, Bash, and/or PowerShell for tooling and automation
  • Deep understanding of networking concepts including TCP/IP, DNS, load balancing, VPNs, and cloud-native networking
  • Experience in SRE, DevOps, or Infrastructure roles, including experience leading engineering teams
  • Proven track record leading incident response and driving reliability improvements

U.S. National Base Pay Range: $118,300 - $219,800. Geographic differentials may apply in some locations to better reflect local market rates.

This job is eligible for an annual incentive bonus.

We know your well-being and happiness are key to a long and successful career. We are delighted to offer country specific benefits.

We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact 1-855-833-5120.

We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.

RELX is a global provider of information-based analytics and decision tools for professional and business customers, enabling them to make better decisions, get better results and be more productive.

Our purpose is to benefit society by developing products that help researchers advance scientific knowledge; doctors and nurses improve the lives of patients; lawyers promote the rule of law and achieve justice and fair results for their clients; businesses and governments prevent fraud; consumers access financial services and get fair prices on insurance; and customers learn about markets and complete transactions.

Our purpose guides our actions beyond the products that we develop. It defines us as a company. Every day across RELX our employees are inspired to undertake initiatives that make unique contributions to society and the communities in which we operate.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX INC • Columbia (SC)

Remote
USD 118.000 - 220.000
Annual incentive bonus
Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX International • Hartford (CT)

Vor Ort
USD 118.000 - 220.000
Annual incentive bonus
Country-specific benefits
Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX International • New York (NY)

Vor Ort
USD 132.000 - 220.000
Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX International • Little Elm (TX)

Vor Ort
USD 118.000 - 220.000
Annual incentive bonus
Country-specific benefits
Site Reliability Engineering Lead
Site Reliability Engineering Lead

RELX International • Tallahassee (FL)

Vor Ort
USD 118.000 - 220.000
Annual incentive bonus
Country-specific benefits
Site Reliability Engineering Lead
Site Reliability Engineering Lead

LexisNexis Risk Solutions • Town of Florida (NY)

Remote
USD 118.000 - 220.000
Annual incentive bonus
Remote-friendly
Site Reliability Engineer II
Site Reliability Engineer II

RELX INC • Pittsburgh

Remote
USD 72.000 - 119.000
Health benefits
401(k) with match
Wellbeing program
+4
Site Reliability Engineer II
Site Reliability Engineer II

RELX INC • Sacramento (CA)

Remote
USD 72.000 - 119.000
Health Benefits
401(k) with match and ESPP
Wellbeing program and Headspace
+3
Site Reliability Engineer
Site Reliability Engineer

RELX INC • USA

Remote
USD 56.000 - 94.000
Senior Software Engineer I
Senior Software Engineer I

RELX INC • Philadelphia

Vor Ort
USD 90.000 - 145.000