Senior SRE: Automation & Reliability Champion

RXinsider LTD.

Philadelphia (Philadelphia County)

On-site

USD 95,000 - 159,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Annual incentive bonus
Country-specific benefits

Job summary

Elsevier is seeking a Senior Site Reliability Engineer to lead reliability initiatives for our critical platforms and services in the United States. You will drive automation, improve observability, and reduce operational toil while ensuring high availability and performance of AI-enabled applications.

You will collaborate with engineering teams to design resilient systems, participate in incident response, and build robust handover capabilities to ensure operability after squads move on.

Qualifications

  • Advanced Terraform including modules, state management, drift detection and remote state handling.
  • Hands-on AWS operations across multi-account, multi-region environments including ECS, RDS, S3, DynamoDB, Lambda, SQS and KMS.
  • Experience with GitHub Actions CI/CD workflows, OIDC auth, and deployment pipelines.
  • Experience with ECS Fargate, Docker, ECR, IAM roles, health checks, autoscaling and deployment rollbacks.
  • Proficiency in AWS Networking & Security: VPCs, ALBs, Route53, TLS, IAM, Secrets Manager, KMS, cloud security best practices.
  • Strong Linux skills and scripting (Bash/Python) for automation and tooling.
  • Experience integrating AI services in production with monitoring and security considerations.
  • Ability to support multiple teams, document solutions, and enable self-service across infra and apps.

Responsibilities

  • Create monitoring queries and establish service level baselines.
  • Support senior engineers during incidents and post-mortems/RCA analysis.
  • Contribute to disaster recovery tests and reliability improvements.
  • Implement automation and execute code in production environments.
  • Document SRE knowledge and runbooks for operations handover.
  • Support deployment, monitoring, and reliability of AI-enabled services.
  • Assist in creating infrastructure topology drawings and deployment workflows.
  • Test availability, reliability, and recoverability in non-production environments.

Skills

Terraform
AWS
CI/CD
Linux
Observability
Incident Response
Python Bash
AI Tooling

Tools

GitHub Actions
ECS Fargate
Docker
ECR
VPC
Route53
CloudWatch

Job description

Elsevier is seeking a Senior Site Reliability Engineer to lead reliability initiatives for our critical platforms and services in the United States. You will drive automation, improve observability, and reduce operational toil while ensuring high availability and performance of AI-enabled applications.

You will collaborate with engineering teams to design resilient systems, participate in incident response, and build robust handover capabilities to ensure operability after squads move on.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE: Reliability, Automation & AI Platforms
Senior SRE: Reliability, Automation & AI Platforms

RX Brasil • Philadelphia

On-site
USD 95,000 - 159,000
Annual incentive bonus
Senior SRE: Automate Reliability & Observability
Senior SRE: Automate Reliability & Observability

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior SRE: Scale Resilient AI Platforms & Automation
Senior SRE: Scale Resilient AI Platforms & Automation

Relx Plc • Philadelphia

Hybrid
USD 95,000 - 159,000
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
Senior SRE Platform Engineer – AI-Powered Reliability
Senior SRE Platform Engineer – AI-Powered Reliability

UiPath • Denver (CO)

Hybrid
USD 160,000 - 210,000
Senior Staff SRE — Reliability, Observability & Automation
Senior Staff SRE — Reliability, Observability & Automation

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 150,000 - 200,000
Healthcare coverage
401(k) with company match
Paid time off & holidays
+2
Senior SRE Lead – Europe, Automation & Observability
Senior SRE Lead – Europe, Automation & Observability

Strike • United States

Remote
USD 102,762 - 194,107
Senior SRE: AI-Driven Reliability & Automation (Hybrid)
Senior SRE: AI-Driven Reliability & Automation (Hybrid)

Namely • United States

Hybrid
USD 120,000 - 150,000
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Senior SRE: Platform Reliability & AI-Driven Ops
Senior SRE: Platform Reliability & AI-Driven Ops

Block • New York (NY)

On-site
USD 170,100 - 283,600
Healthcare coverage
Retirement plans
Employee Stock Purchase Program
+1