Senior SRE (AWS)

VIQU Ltd

Milton Keynes

On-site

GBP 60,000 - 75,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Bonus
On-call allowance
On-site 2 days/week

Job summary

VIQU Ltd is seeking a Senior Site Reliability Engineer to own cloud stability and incident response for a growing B2B SaaS platform. You will drive automation, observability, and tooling adoption while collaborating with cross-functional teams.

You will work primarily with AWS and on-premise VMs, implementing IaC with Terraform, Kubernetes, and monitoring stacks such as Prometheus, Grafana, and Datadog. On-site presence is required two days per week in Milton Keynes.

Qualifications

  • Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment - e.g. SaaS or MSP.
  • Strong hands-on experience with both AWS, and on-premise virtual machines.
  • Experience with Infrastructure as Code/Terraform, Container orchestration (Kubernetes), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor).
  • Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working.
  • Ability to communicate across internal teams and external customers.
  • Skilled in networking across both cloud (Azure) and on premise environments.
  • Either Windows or Linux systems administration skills (Linux preferred).
  • Previous use of AI tools to enhance efficiency.

Responsibilities

  • Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure Servers and networks, and automate application life cycles.
  • Regularly use Datadog and other observability tools for application performance monitoring.
  • Implement new ways of working, helping to shape how the organisation responds and recovers to incidents.
  • Take ownership of incident resolutions.
  • Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents.
  • Work on an a on call rota, ensuring you are available to respond to incidents during this time.
  • Identify areas for automation and help implement changes that raise the bar for reliability.

Skills

Cloud & DevOps
Incident response
Cross-team communication
On-call rotation
Cloud networking
Linux/Windows admin
AI tooling usage

Tools

AWS
Terraform
Kubernetes
Prometheus
Grafana
Datadog
Azure Monitor

Job description

Senior Site Reliability Engineer (AWS focused)

Up to £75,000 + bonus + on call allowance

Milton Keynes (2 days on site a week)

VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation.

This is a genuine opportunity to own and operate how the cloud function works, and progress into a team lead position as the team grows.

Experience required for the Senior Site Reliability Engineer
  • Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment - eg SaaS or MSP.
  • Strong hands-on experience with both AWS, and on-premise virtual machines.
  • Experience withInfrastructure as Code/Terraform, Container orchestration (Kubernetes), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor).
  • Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working.
  • Ability to communicate across internal teams and external customers.
  • Skilled in networking across both cloud (Azure) and on premise environments.
  • Either Windows or Linux systems administration skills (Linux preferred).
  • Previous use of AI tools to enhance efficiency.
Job Duties of the Senior Site Reliability Engineer
  • Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure Servers and networks, and automate application life cycles.
  • Regularly use Datadog and other observability tools for application performance monitoring.
  • Implement new ways of working, helping to shape how the organisation responds and recovers to incidents.
  • Take ownership of incident resolutions.
  • Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents.
  • Work on an a on call rota, ensuring you are available to respond to incidents during this time.
  • Identify areas for automation and help implement changes that raise the bar for reliability.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE (AWS)
Senior SRE (AWS)

VIQU IT • Milton Keynes

On-site
GBP 68,000 - 83,000
On-site 2 days/week
Bonus
On-call allowance
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VIQU IT Recruitment • Milton Keynes

On-site
GBP 45,000 - 75,000
Senior AWS SRE: Platform Reliability & Automation
Senior AWS SRE: Platform Reliability & Automation

VIQU IT • Milton Keynes

On-site
GBP 68,000 - 83,000
On-site 2 days/week
Bonus
On-call allowance
+1
Senior AWS SRE | Hybrid On-Site, Reliability Leader
Senior AWS SRE | Hybrid On-Site, Reliability Leader

VIQU Ltd • Milton Keynes

On-site
GBP 60,000 - 75,000
Bonus
On-call allowance
On-site 2 days/week
Senior Site Reliability Engineer — Reliability Lead
Senior Site Reliability Engineer — Reliability Lead

VIQU IT Recruitment • Milton Keynes

On-site
GBP 45,000 - 75,000
Senior Site Reliability Engineer (LON)
Senior Site Reliability Engineer (LON)

McNally Recruitment Ltd • Greater London

Hybrid
GBP 90,000 - 150,000
Benefits as Cash
Hybrid work model
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Spectrum IT Recruitment • Southampton

Hybrid
GBP 55,000 - 90,000
Life Insurance
Private Medical Insurance
Employee Assistance Programme
+3
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Spectrum IT • Southampton

Hybrid
GBP 90,000 - 120,000
Senior Site Reliability Engineer (Application / API Focused)
Senior Site Reliability Engineer (Application / API Focused)

Xpertise Recruitment • Greater London

On-site
GBP 90,000 - 110,000
25% Bonus
Excellent Benefits
Site Reliability Engineer
Site Reliability Engineer

ReVybe IT Recruitment Limited • Greater London

On-site
GBP 51,000 - 85,000
Bonus
Benefits