Site Reliability Engineer

Atarus

Greater London

On-site

GBP 90,000 - 150,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Atarus partners with a fast-growing UK AI company to hire a Senior Site Reliability Engineer. You will own reliability across both the AWS platform and edge infrastructure, combining software engineering, automation and production infrastructure at scale.

You will build tooling, manage large-scale AWS and Kubernetes environments, improve observability and CI/CD, and automate provisioning for thousands of devices while enhancing developer productivity through internal tooling.

Qualifications

  • Strong SRE or software engineering background.
  • Production coding experience with Python or Go.
  • Infrastructure as Code and automation expertise.
  • Understanding of observability and incident management.
  • Edge/IoT experience is a bonus, not essential.
  • Role involves coding as part of SRE responsibilities.
  • Experience spanning cloud and thousands of devices.
  • Exposure to AI-scale systems and high availability.

Responsibilities

  • Build software and tooling to improve reliability and scalability of production systems.
  • Own and evolve large-scale AWS & Kubernetes environments.
  • Build observability, monitoring and telemetry across cloud and edge infrastructure.
  • Improve deployment, CI/CD and incident response processes.
  • Automate provisioning and lifecycle management across connected devices.
  • Identify reliability bottlenecks and engineer them out of the platform.
  • Improve developer experience through internal tooling and automation.

Skills

SRE background
Python/Go
Infrastructure as Code
Observability
Cloud + Edge
Software writing
IoT/Edge experience
Production reliability
Automation
Kubernetes

Job description

Senior Site Reliability Engineer (SRE)

We’re partnered with one of the UK’s fastest-growing AI companies, building cloud-native video intelligence products deployed across a rapidly growing fleet of connected devices.


They’re looking for a Senior SRE to own reliability across both their AWS platform and edge infrastructure — combining software engineering, automation and production infrastructure at serious scale.


What You’ll Be Doing


  • Build software and tooling to improve the reliability and scalability of production systems

  • Own and evolve large-scale AWS & Kubernetes environments

  • Build observability, monitoring and telemetry across cloud and edge infrastructure

  • Improve deployment, CI/CD and incident response processes

  • Automate provisioning and lifecycle management across connected devices

  • Identify reliability bottlenecks and engineer them out of the platform

  • Improve developer experience through internal tooling and automation


What They’re Looking For


  • Strong SRE or software engineering background

  • Production coding experience with Python or Go

  • Strong Infrastructure-as-Code and automation experience

  • Solid understanding of observability, incident management and production reliability

  • Edge / IoT experience is a bonus, not essential

  • SRE role where you'll actually write software, not just manage infrastructure

  • Reliability challenges spanning cloud + thousands of physical devices

  • Real-world AI systems with meaningful scale and availability requirements

  • Small engineering team with significant ownership and architectural influence

  • Strong commercial traction, significant funding and meaningful equity


If you’re an SRE who enjoys building systems rather than babysitting them, this is worth a conversation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE
Senior SRE

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Xpertise Recruitment • West Drayton

On-site
GBP 60,000 - 80,000
Site Reliability Engineer
Site Reliability Engineer

DNS INFO LTD • City Of London

On-site
GBP 70,000 - 95,000
SRE
SRE

SR2 REC LTD • Greater London

On-site
GBP 70,000 - 110,000
Senior Site Reliability Engineer (Application / API Focused)
Senior Site Reliability Engineer (Application / API Focused)

Xpertise Recruitment • Greater London

Hybrid
GBP 80,000 - 100,000
25% Bonus
Excellent Benefits
Site Reliability Engineer
Site Reliability Engineer

Incite-Insight.co.uk • West of England

On-site
GBP 70,000 - 95,000
Senior SRE: Cloud & Edge Reliability Engineer
Senior SRE: Cloud & Edge Reliability Engineer

Atarus • Greater London

On-site
GBP 90,000 - 150,000
SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
SRE Technical Lead
SRE Technical Lead

83zero Ltd • Wokingham

Hybrid
GBP 60,000 - 100,000
5% bonus
Hybrid working model
Site Reliability Engineer
Site Reliability Engineer

Wedo Technology Solutions Ltd. • Greater London

Remote
GBP 63,000 - 75,000