Senior SRE: Scale Resilient AI Platforms & Automation

Relx Plc

Philadelphia (Philadelphia County)

Hybrid

USD 95,000 - 159,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Relx Plc seeks a Senior Site Reliability Engineer to design and operate scalable, resilient platforms powering mission-critical apps. You will lead reliability initiatives, automate toil, and collaborate with engineering teams to boost availability and performance.

You will use observability, incident response, and AI tooling deployment; manage infrastructure with Terraform on AWS, and ensure secure, well-documented handover as squads evolve.

Qualifications

  • Advanced Terraform expertise with modules, providers, state management and drift detection.
  • Hands-on AWS operations across multi-account, multi-region environments.
  • Experience with CI/CD using GitHub Actions and Terraform deployments.
  • Knowledge of ECS Fargate, Docker, ECR, health checks, autoscaling and deployments.
  • Networking and security in AWS (VPC, Route53, TLS/IAM, KMS).
  • Strong incident response, observability, root-cause analysis and runbooks.
  • Linux scripting (Bash/Python) for automation and AWS CLI workflows.
  • Experience deploying AI tools and services in production with monitoring.
  • Ability to enable secure self-service and support multiple teams.

Responsibilities

  • Create monitoring queries and establish service level baselines.
  • Support incidents and contribute to post-mortems and RCAs.
  • Participate in disaster recovery tests and automation efforts.
  • Develop deployment workflows and maintain reliable infrastructure topology drawings.
  • Document SRE knowledge and support AI tooling deployments in prod.
  • Ensure availability, recoverability, and test non-production environments.

Skills

Advanced Terraform
AWS Operations
GitHub Actions
ECS Fargate
AWS Networking
Incident Response
Linux Automation
AI Tooling Deployment
Developer Enablement

Tools

Terraform
AWS
Docker
ECS
ECR
S3
Route53
CloudWatch
GitHub Actions

Job description

Relx Plc seeks a Senior Site Reliability Engineer to design and operate scalable, resilient platforms powering mission-critical apps. You will lead reliability initiatives, automate toil, and collaborate with engineering teams to boost availability and performance.

You will use observability, incident response, and AI tooling deployment; manage infrastructure with Terraform on AWS, and ensure secure, well-documented handover as squads evolve.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE: AI-Driven Cloud Reliability & Automation
Senior SRE: AI-Driven Cloud Reliability & Automation

Hidden Jobs • United States

Remote
USD 191,000 - 226,000
Equity incentive
Flexible PTO
Health insurance
+2
Senior SRE: Scale systems, automate with IaC
Senior SRE: Scale systems, automate with IaC

Replit • United States

Remote
USD 120,000 - 210,000
Competitive Salary
Equity
401(k) Match
+16
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
Senior SRE: AI-Driven Platform Reliability & Scale
Senior SRE: AI-Driven Platform Reliability & Scale

Medallia • McLean (VA)

On-site
USD 129,000 - 190,000
Health benefits
401(k) matching
Paid parental leave
+1
Senior SRE II — Scale Systems with AI-Driven Reliability
Senior SRE II — Scale Systems with AI-Driven Reliability

Juniper Square • United States

On-site
USD 165,000 - 195,000
Health, dental, and vision care
Life insurance
Mental wellness coverage
+3
Senior SRE Platform Engineer – AI-Powered Reliability
Senior SRE Platform Engineer – AI-Powered Reliability

UiPath • Denver (CO)

Hybrid
USD 160,000 - 210,000
Senior SRE: Scale Reliability, Observability & Resilience
Senior SRE: Scale Reliability, Observability & Resilience

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Senior SRE — AI Resilience & Cloud Platforms
Senior SRE — AI Resilience & Cloud Platforms

Cisco • San Jose (CA)

Hybrid
USD 168,000 - 245,000
Senior SRE: AI-Driven Reliability & Cloud Automation
Senior SRE: AI-Driven Reliability & Cloud Automation

NDEAVOUR CONSULTING • United States

Hybrid
USD 120,000 - 150,000
Remote Office
Parking Space
Fun Office Space
+7
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform
Senior SRE: Scale & Reliability for AI-Driven SaaS Platform

Instrumental Inc. • Palo Alto (CA)

On-site
USD 175,000 - 229,000
Health benefits
Commuter plans
Parental leave