Senior SRE — Automation & Reliability Leader

Elsevier

Greater London

In loco

GBP 95.000 - 130.000

Tempo pieno

11 giorni fa
Generatore di candidature

Una candidatura fatta su misura per questo lavoro — un curriculum e una lettera di presentazione personalizzati, perfettamente in linea con l'annuncio.

Supera i filtri ATS

Vantaggi offerti da questo lavoro

Comprehensive Pension Plan
Generous vacation entitlement
Sabbatical leave option
Employee discounts
Family care leave

Descrizione del lavoro

Elsevier seeks a Senior Site Reliability Engineer to ensure reliability, scalability, and performance of critical platforms. You will lead complex reliability initiatives, automate operations, and build resilient systems used across AI-powered services.

You will leverage observability, incident response, and distributed systems expertise to design solutions that improve availability, streamline ops, and enable secure self-service practices for multiple teams.

Competenze

  • Advanced Terraform capabilities including modules, providers, state management and remote state handling.
  • Hands-on AWS operations across multi-account/multi-region environments (ECS, RDS, S3, Lambda, DynamoDB, etc.).
  • Experience building/reapplying CI/CD pipelines with GitHub Actions and Terraform.
  • Strong knowledge of ECS Fargate, containers, IAM roles, health checks, autoscaling and rollback strategies.
  • Proficiency in AWS networking and security best practices (VPC, Route53, ACM/TLS, KMS, Secrets Manager).
  • Incident response, observability, and capability to perform RCAs and post-mortems.
  • Strong Linux scripting (Bash/Python) for AWS CLI automation and tooling.
  • Experience deploying AI tooling in production with monitoring and security considerations.
  • Ability to enable multiple engineering teams with secure self-service practices.

Mansioni

  • Create monitoring queries and establish service baselines.
  • Support incidents and contribute to post-mortems and RCAs.
  • Participate in disaster recovery tests and automation initiatives.
  • Implement automation and run code in production environments.
  • Contribute to SRE knowledge documentation and playbooks.
  • Support deployment, monitoring, and reliability of AI-integrated services.
  • Provide direction on platform features and CI/CD improvements.

Conoscenze

Terraform
AWS
GitHub Actions
ECS Fargate & Containers
Networking & Security
Incident Response & Observability
Linux & Automation
AI Tooling Deployment
Developer Enablement

Descrizione del lavoro

Elsevier seeks a Senior Site Reliability Engineer to ensure reliability, scalability, and performance of critical platforms. You will lead complex reliability initiatives, automate operations, and build resilient systems used across AI-powered services.

You will leverage observability, incident response, and distributed systems expertise to design solutions that improve availability, streamline ops, and enable secure self-service practices for multiple teams.

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Senior SRE: Build Resilient, AI-Driven Systems
Senior SRE: Build Resilient, AI-Driven Systems

Elsevier • City Of London

In loco
GBP 70.000 - 120.000
Senior SRE - Reliability & Automation Lead (Flexible Hours)
Senior SRE - Reliability & Automation Lead (Flexible Hours)

Elsevier Limited Company • City Of London

In loco
GBP 90.000 - 120.000
Comprehensive Pension Plan
Generous vacation + sabbatical leave
Family leave (Maternity, Paternity, Ad
+5
Senior SRE: Automate, Scale & Improve Reliability
Senior SRE: Automate, Scale & Improve Reliability

RELX International • Greater London

In loco
GBP 90.000 - 120.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

RELX International • Greater London

In loco
GBP 90.000 - 120.000
Senior SRE for AI Platform & HPC
Senior SRE for AI Platform & HPC

Mistral Ai • York and North Yorkshire

Ibrido
GBP 70.000 - 110.000
Healthcare coverage
Relocation support
Wellness programs
Senior SRE — Build a Global Cloud Reliability Practice
Senior SRE — Build a Global Cloud Reliability Practice

Omnicell • Manchester

Ibrido
GBP 90.000 - 130.000
SRE Director — AI-Driven Reliability & Scale
SRE Director — AI-Driven Reliability & Scale

EPAM Systems • Greater London

Ibrido
GBP 180.000 - 240.000
ESPP
Life Assurance
Income protection
+14
Senior SRE: Cloud Reliability & Observability
Senior SRE: Cloud Reliability & Observability

Renesas Electronics Corp. • Cambridge

In loco
GBP 90.000 - 120.000
SRE Manager: Scale, Reliability & Observability Leader
SRE Manager: Scale, Reliability & Observability Leader

UST • Nottingham

In loco
GBP 90.000 - 120.000
Senior SRE, AWS Platform & Reliability Lead
Senior SRE, AWS Platform & Reliability Lead

Aumni • Glasgow

In loco
GBP 90.000 - 130.000