Site Reliability Engineering Lead

LexisNexis Risk Solutions

Town of Florida (NY)

Remote

USD 118,000 - 220,000

Full time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Annual incentive bonus
Remote-friendly

Job summary

LexisNexis Risk Solutions is seeking a Technology Site Reliability Engineering Lead to coach a small to mid-size SRE team and own reliability initiatives across cloud environments. You will drive automation, cost optimization, incident response, and post-mortems while shaping CI/CD and platform resilience.

The role requires expert Kubernetes, Terraform, Azure skills, and strong leadership to mentor engineers, plan capacity, and collaborate with safety, product, and Dev teams.

Qualifications

  • Strong leadership and people management experience in SRE/DevOps.
  • Expert Kubernetes and cloud native architecture skills.
  • Hands-on Terraform and IaC practices.
  • Proficient in GitHub Actions for CI/CD and release automation.
  • Experience with observability using Prometheus, Grafana, OpenTelemetry.
  • Proficient in Python, Bash and/or PowerShell scripting.
  • Deep understanding of networking, DNS, load balancing, and VPNs.
  • Proven track record leading incident response and reliability improvements.

Responsibilities

  • Manage, mentor, and grow a team of SREs; conduct 1:1s and performance reviews.
  • Own hiring, onboarding, and team capacity decisions.
  • Set team goals, backlog, and drive planning.
  • Foster blameless post-incident culture and cross-team collaboration.
  • Lead reliability initiatives across infrastructure and services.
  • Drive incident response and continuous service improvement.
  • Champion automation and operational excellence across the platform.
  • Support scalable, secure, and resilient cloud-native environments.

Skills

Team leadership
Incident response
Automation
Python
Bash/PowerShell
Networking concepts
SRE/DevOps
Cost optimization
Cloud concepts
Post-mortems

Tools

Kubernetes
Terraform
Azure Cloud
GitHub Actions
Prometheus
Grafana
OpenTelemetry
CI/CD
IaC

Job description

Technology
Site Reliability Engineering Lead

Are you passionate about building reliable, scalable platforms and helping engineering teams thrive?

Do you enjoy leading high-performing teams while driving automation, resilience, and operational excellence across modern cloud environments?

About the Business

LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Business Services vertical, we offer a multitude of solutions focused on helping businesses of all sizes drive higher revenue growth, maximize operational efficiencies, and improve customer experience. Our solutions help our customers solve difficult problems in the areas of Anti-Money Laundering/Counter Terrorist Financing, Identity Authentication & Verification, Fraud and Credit Risk mitigation and Customer Data Management. You can learn more about LexisNexis Risk athttps://risk.lexisnexis.com/

About Our Team

You will be joining the Core SRE Team in Business Services, a team that oversees all the applications and infrastructure in the biggest business unit in LexisNexis Risk Solutions. We build cloud environments, migrate on prem applications to the cloud, work with self-hosted and 3rdparty solutions. The successful candidate is a self-starter who assesses the situation, collaboratively develops a solution, and takes the initiative to improve performance, cost, and reliability at each opportunity.

About the Role

This is a professional management level role. Individuals are required to provide line management to a small to medium sized team of engineers, including authority over performance management, pay, and recruitment. They will ensure that tasks and projects are prioritized appropriately and provide support to team members when tasks are blocked. They will support engineers in their personal development and ensure they are working within the SRE framework. They will lead the post-mortem reviews and the timely production of RCAs. They address issues with impact beyond their own team based on knowledge of related disciplines.

Responsibilities

  • Manage, mentor, and grow a team of SREs; conduct 1:1s, performance reviews, and career development planning
  • Own hiring, onboarding, and team capacity/resourcing decisions
  • Set team goals, prioritize backlog, and drive planning
  • Foster a blameless post-incident culture and cross-team collaboration with Dev, Security, and Product
  • Lead reliability initiatives across infrastructure and services
  • Drive incident response activities and continuous service improvement
  • Champion automation and operational excellence across the platform
  • Support the development of scalable, secure, and resilient cloud-native environments

Requirements

  • Expert knowledge of Kubernetes, including cluster architecture, upgrades, autoscaling, security hardening, and troubleshooting at scale
  • Expert experience with Terraform, including modular IaC design, state management, multi-environment provisioning, and policy-as-code
  • Deep knowledge of Azure Cloud, including compute, networking, identity (AAD), storage, and cost optimization
  • Experience designing and scaling CI/CD pipelines using GitHub Actions, release strategies, and rollback automation
  • Experience with observability platforms including Prometheus, Grafana, OpenTelemetry, and SLO/SLA/error-budget management
  • Strong automation skills, focused on eliminating toil through self-healing systems and infrastructure automation
  • Advanced proficiency in Python, Bash, and/or PowerShell for tooling and automation
  • Deep understanding of networking concepts including TCP/IP, DNS, load balancing, VPNs, and cloud-native networking
  • Experience in SRE, DevOps, or Infrastructure roles, including experience leading engineering teams
  • Proven track record leading incident response and driving reliability improvements

Additional location(s): Home based-Connecticut; Home based-New Jersey; Home based-New York; Home based-Pennsylvania; Home based-South Carolina; Home based-Delaware; Home based-Washington DC

U.S. National Base Pay Range: $118,300 - $219,800. Geographic differentials may apply in some locations to better reflect local market rates.
This job is eligible for an annual incentive bonus.

We are committed to providing a fair and accessible hiring process. If you have a disability or other need that requires accommodation or adjustment, please let us know by completing our Applicant Request Support Form or please contact 1-855-833-5120.

We are an equal opportunity employer: qualified applicants are considered for and treated during employment without regard to race, color, creed, religion, sex, national origin, citizenship status, disability status, protected veteran status, age, marital status, sexual orientation, gender identity, genetic information, or any other characteristic protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineering Lead
Site Reliability Engineering Lead

Relx Plc • York

Remote
USD 144,000 - 195,000
Annual incentive bonus
Country-specific benefits
Site Reliability Engineer II
Site Reliability Engineer II

Talentify • Pittsburgh

On-site
USD 72,000 - 119,000
Health benefits
401(k) with match
Wellbeing program
+4
Site Reliability Engineering Lead
Site Reliability Engineering Lead

Talent Octopusventures • United States

Hybrid
USD 118,000 - 220,000
Medical Insurance
Life Assurance Policies
Flexible Benefits Plan
+3
Site Reliability Engineering Lead
Site Reliability Engineering Lead

LexisNexis Risk Solutions • Alpharetta (GA)

On-site
USD 118,000 - 220,000
Site Reliability Engineering Lead
Site Reliability Engineering Lead

LexisNexis Risk Solutions • New York (NY)

On-site
USD 142,000 - 264,000
Medical Insurance
Life Insurance
Flexible Benefits Plan
+4
Site Reliability Engineering Lead
Site Reliability Engineering Lead

LexisNexis Special Services Inc. • Alpharetta (GA), Northern (KY)

Hybrid
USD 118,000 - 264,000
Annual incentive bonus
Software Engineering Lead
Software Engineering Lead

RELX International • United States

Remote
USD 115,000 - 192,000
Senior Site Reliability Engineer II
Senior Site Reliability Engineer II

LexisNexis Risk Solutions FL Inc. Company • San Jose (CA)

On-site
USD 105,000 - 175,000
Health benefits
401(k) with match
Employee Wellness programs
+1
Software Engineering Lead
Software Engineering Lead

LexisNexis Risk Solutions • Alpharetta (GA), Northern (KY)

On-site
USD 115,000 - 192,000
Site Reliability Engineer
Site Reliability Engineer

RELX INC • United States

Remote
USD 56,000 - 94,000