Senior SRE - Cloud & Edge Reliability Engineer

Atarus

England

On-site

GBP 110,000 - 170,000

Full time

38 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Atarus is partnering with a fast-growing UK AI company to recruit a Senior Site Reliability Engineer. You’ll own reliability across AWS cloud and edge infrastructure, combining software engineering, automation and production infrastructure at scale.

Build tooling to improve reliability, own large-scale AWS & Kubernetes environments, and enhance observability, deployment pipelines and incident response. We value production coding in Python or Go, IaC expertise, incident management and a bias for

Qualifications

  • Strong SRE or software engineering background.
  • Production coding experience with Python or Go.
  • Infrastructure-as-Code and automation experience.
  • Observability, incident management and production reliability.
  • Edge / IoT experience is a bonus, not essential.
  • SRE role where you'll actually write software, not just manage infrastructure.
  • Reliability challenges spanning cloud + thousands of physical devices.
  • Real-world AI systems with meaningful scale and availability requirements.
  • Small engineering team with significant ownership and architectural influence.
  • Strong commercial traction, significant funding and meaningful equity.

Responsibilities

  • Build software and tooling to improve the reliability and scalability of production systems.
  • Own and evolve large-scale AWS & Kubernetes environments.
  • Build observability, monitoring and telemetry across cloud and edge infrastructure.
  • Improve deployment, CI/CD and incident response processes.
  • Automate provisioning and lifecycle management across connected devices.
  • Identify reliability bottlenecks and engineer them out of the platform.
  • Improve developer experience through internal tooling and automation.

Skills

Python
Go
Infrastructure as Code
Observability
Automation
CI/CD
AWS
Kubernetes

Tools

Terraform
Docker

Job description

Atarus is partnering with a fast-growing UK AI company to recruit a Senior Site Reliability Engineer. You’ll own reliability across AWS cloud and edge infrastructure, combining software engineering, automation and production infrastructure at scale.

Build tooling to improve reliability, own large-scale AWS & Kubernetes environments, and enhance observability, deployment pipelines and incident response. We value production coding in Python or Go, IaC expertise, incident management and a bias for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AWS SRE: Build Resilient Cloud Platforms
Remote AWS SRE: Build Resilient Cloud Platforms

SPECTRUM IT • England

On-site
GBP 70,000 - 110,000
Fully remote (UK)
24/7 shift pattern
Bonus & benefits
Remote Senior Azure SRE - Reliability & Automation
Remote Senior Azure SRE - Reliability & Automation

OMEGA, Inc. • United Kingdom

Remote
GBP 100,000 - 140,000
Competitive salary and benefits
Professional development
Certification support
+1
Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Senior SRE: Cloud Reliability & Observability Leader
Senior SRE: Cloud Reliability & Observability Leader

Omilia • United Kingdom

On-site
GBP 90,000 - 130,000
Fixed compensation
Long-term employment
Professional growth
+1
Edge & Cloud Infrastructure Engineer | AI Platform DevOps
Edge & Cloud Infrastructure Engineer | AI Platform DevOps

Atarus • England

On-site
GBP 70,000 - 110,000
Senior SRE: Cloud, CI/CD & Scalable Infra — London
Senior SRE: Cloud, CI/CD & Scalable Infra — London

Durlston Partners • Greater London

On-site
GBP 110,000 - 150,000
Senior SRE: Cloud Reliability, Observability & Automation
Senior SRE: Cloud Reliability, Observability & Automation

Omilia Natural Language Solutions Ltd • United Kingdom

On-site
GBP 70,000 - 90,000
Fixed compensation
Long-term employment with vacation
Professional growth opportunities
+1
Senior SRE - AWS, Kubernetes & Observability
Senior SRE - AWS, Kubernetes & Observability

ReVybe IT Recruitment Limited • Greater London

Hybrid
GBP 51,000 - 85,000
Bonus
Benefits
AWS SRE: Cloud Data Platform Reliability & Automation
AWS SRE: Cloud Data Platform Reliability & Automation

Marks Sattin • Greater London

On-site
GBP 60,000 - 80,000
Senior SRE: Build Fault-Tolerant AI Cloud Infra
Senior SRE: Build Fault-Tolerant AI Cloud Infra

Nebius • Greater London

On-site
GBP 90,000 - 120,000
Competitive pay
Career growth
Flexible work
+3