Senior Platform SRE: Multi-Cloud Reliability & Resilience

Elastic

Greater London

Hybrid

GBP 90,000 - 130,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Health coverage for you and family
Flexible location & schedule
Generous vacation days
16+ weeks parental leave
Volunteer time
Charitable giving matched

Job summary

Elastic is seeking an experienced Site Reliability Engineer within the Platform Engineering team to design, build, and scale a multi-cloud platform that hosts Elastic Cloud Hosted and Serverless. You will automate system engineering tasks to guarantee global reliability and participate in follow-the-sun on-call rotations.

You will work with Golang, Kubernetes, and IaC tools to extend tooling, monitor performance, and improve incident response across distributed teams and environments.

Qualifications

  • Experience building and operating SaaS platforms in public cloud environments.
  • Strong programming skills in Golang or similar languages.
  • Hands-on experience with Kubernetes across multiple clouds.
  • Expertise in alerting and major incident management practices.
  • Familiarity with IaC tooling (Terraform, Crossplane) and containerized services.

Responsibilities

  • Lead automation to guarantee reliability of Elastic infrastructure.
  • Participate in on-call rotations with follow-the-sun coverage.
  • Respond to incidents, perform root cause analyses and implement durable fixes.
  • Design and scale multi-cloud platform components for Elastic Cloud Hosted and Serverless.
  • Collaborate across distributed teams to elevate reliability and performance.

Skills

Golang
Kubernetes
Terraform
Crossplane
Linux
Public cloud
Incident management
Alerting
On-call
Distributed teams

Tools

Elastic Stack
Prometheus
Graphite
Influx
Docker

Job description

Elastic is seeking an experienced Site Reliability Engineer within the Platform Engineering team to design, build, and scale a multi-cloud platform that hosts Elastic Cloud Hosted and Serverless. You will automate system engineering tasks to guarantee global reliability and participate in follow-the-sun on-call rotations.

You will work with Golang, Kubernetes, and IaC tools to extend tooling, monitor performance, and improve incident response across distributed teams and environments.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud SRE: Multi-Cloud Reliability & Automation
Senior Cloud SRE: Multi-Cloud Reliability & Automation

Elasticsearch B.V. • United Kingdom

Remote
GBP 75,000 - 140,000
Health coverage
Flexible locations and schedules
Generous vacation days
+4
Senior Site Reliability Engineer (Platform Reliability, Resilience)
Senior Site Reliability Engineer (Platform Reliability, Resilience)

Elastic • Greater London

Hybrid
GBP 90,000 - 130,000
Health coverage for you and family
Flexible location & schedule
Generous vacation days
+3
Remote Cloud Site Reliability Manager
Remote Cloud Site Reliability Manager

Cameroon Mathematical Union (CAMU) • Greater London

Hybrid
GBP 15,000 - 38,000
Remote AWS SRE: Build Resilient Cloud Platforms
Remote AWS SRE: Build Resilient Cloud Platforms

SPECTRUM IT • England

On-site
GBP 70,000 - 110,000
Fully remote (UK)
24/7 shift pattern
Bonus & benefits
Senior AI Platform SRE: Reliability & Automation
Senior AI Platform SRE: Reliability & Automation

CloudFactory • Reading

On-site
GBP 70,000 - 110,000
Principal Cloud SRE: Scale & Reliability (AWS/Azure)
Principal Cloud SRE: Scale & Reliability (AWS/Azure)

LSEG • Greater London

On-site
GBP 75,000 - 100,000
Healthcare
Retirement planning
Paid volunteering days
+1
Senior SRE — Build a Global Cloud Reliability Practice
Senior SRE — Build a Global Cloud Reliability Practice

Omnicell • Manchester

Hybrid
GBP 90,000 - 130,000
Senior Cloud SRE for AI Platform — Reliability & Scale
Senior Cloud SRE for AI Platform — Reliability & Scale

Mistral • Greater London

On-site
GBP 90,000 - 140,000
Healthcare coverage
Parental leave
Retirement plans
+3
Senior Platform & SRE Leader — Remote
Senior Platform & SRE Leader — Remote

aitrainer • United Kingdom

Remote
GBP 120,000 - 180,000
Cloud SRE — Build Scalable Infra & Kubernetes Platforms
Cloud SRE — Build Scalable Infra & Kubernetes Platforms

LexisNexis Risk Solutions • England

Remote
GBP 70,000 - 95,000