Senior SRE - Reliability, Cloud & Kubernetes Lead

RELX International

San Jose (CA)

On-site

USD 105,000 - 175,000

Full time

13 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health benefits
401(k) with match
Wellbeing program
Disability & life insurance
Family benefits
Commuter benefits
Volunteer leave

Job summary

RELX International in San Jose, CA is seeking a Senior Site Reliability Engineer to improve reliability, availability, and performance of production systems across multiple environments.

You will lead incident responses, postmortems, root-cause analyses, and remediation planning while collaborating with engineering, security, and operations teams to ensure timely resolution of incidents and changes.

Qualifications

  • 5+ years of experience in Site Reliability Engineering, Systems Engineering, DevOps, Infrastructure Engineering, or a related field.
  • Bachelor’s degree in Engineering, Computer Science, Information Technology, or equivalent professional experience.
  • Demonstrated experience supporting highly available production systems.
  • Experience leading incident reviews, postmortems, root-cause analysis, and remediation planning.
  • Experience working across infrastructure, application, security, and operations teams.
  • Strong problem-solving, analytical, organizational, and communication skills.
  • Ability to manage multiple priorities and drive work to completion in a fast-paced operational environment.

Responsibilities

  • Lead and participate in incident response, postmortems, root-cause analysis, and gap assessments.
  • Identify reliability, availability, performance, security, and operational risks across production environments.
  • Develop, prioritize, and track corrective and preventive actions through completion.
  • Follow up with engineering, development, security, support, and business stakeholders to ensure timely resolution of incidents and identified gaps.
  • Respond to system-management alerts and operational exceptions within assigned enterprise systems and product offerings.
  • Provide technical input into project plans, schedules, implementation methodologies, and operational readiness activities.
  • Support the triage, planning, execution, documentation, and closure of changes, service requests, and operational tasks.
  • Lead or contribute to Operations Team projects involving cloud, on-premises infrastructure, security, Kubernetes, automation, monitoring, and system modernization.
  • Improve production quality and availability by creating new operational capabilities and remediating weaknesses in existing systems and processes.

Skills

SRE
Cloud
Kubernetes
Automation
IaC
Python
Shell/PowerShell
Incident response

Education

Bachelor’s degree in Engineering/Computer Science/Information Technology

Job description

RELX International in San Jose, CA is seeking a Senior Site Reliability Engineer to improve reliability, availability, and performance of production systems across multiple environments.

You will lead incident responses, postmortems, root-cause analyses, and remediation planning while collaborating with engineering, security, and operations teams to ensure timely resolution of incidents and changes.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE II: Drive Reliability & Incident Mastery
Senior SRE II: Drive Reliability & Incident Mastery

LexisNexis Risk Solutions • San Jose (CA), Northern (KY)

Hybrid
USD 105,000 - 175,000
401(k) with match
Wellbeing programs
Life Insurance
+1
Senior SRE II: Reliability, Automation & Incident Champion
Senior SRE II: Reliability, Automation & Incident Champion

LexisNexis Risk Solutions FL Inc. Company • San Jose (CA)

On-site
USD 105,000 - 175,000
Health benefits
401(k) with match
Employee Wellness programs
+1
SRE Lead: Cloud Reliability & Automation Leader
SRE Lead: Cloud Reliability & Automation Leader

LexisNexis Risk Solutions • Town of Florida (NY)

Remote
USD 118,000 - 220,000
Annual incentive bonus
Remote-friendly
SRE Lead: Drive Reliability, Cloud Modernization
SRE Lead: Drive Reliability, Cloud Modernization

LexisNexis Risk Solutions • Alpharetta (GA)

On-site
USD 118,000 - 220,000
SRE Lead: Drive Reliability & Automation
SRE Lead: Drive Reliability & Automation

Relx Plc • York

Remote
USD 144,000 - 195,000
Annual incentive bonus
Country-specific benefits
Remote Senior Network Reliability Engineer (SRE)
Remote Senior Network Reliability Engineer (SRE)

Gainbridge • Zionsville (IN), Northern (KY)

On-site
USD 135,000 - 190,000
Health Insurance
Dental Insurance
Vision Insurance
+4
Remote SRE Lead: Cloud Modernization & Reliability
Remote SRE Lead: Cloud Modernization & Reliability

LexisNexis Special Services Inc. • Alpharetta (GA), Northern (KY)

Hybrid
USD 118,000 - 264,000
Annual incentive bonus
Lead SRE: Cloud, Automation & Reliability
Lead SRE: Cloud, Automation & Reliability

Federal Reserve Bank of San Francisco • Dallas (TX)

On-site
USD 147,000 - 234,000
Senior SRE - Observability & Cloud Reliability
Senior SRE - Observability & Cloud Reliability

Cisco Systems, Inc. • San Francisco (CA)

On-site
USD 168,000 - 245,000
Medical benefits
401(k) matching
Parental leave
+1
Senior SRE: Scale Reliability, Observability & Resilience
Senior SRE: Scale Reliability, Observability & Resilience

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2