Senior AWS SRE | Hybrid On-Site, Reliability Leader

VIQU Ltd

Milton Keynes

On-site

GBP 60,000 - 75,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Bonus
On-call allowance
On-site 2 days/week

Job summary

VIQU Ltd is seeking a Senior Site Reliability Engineer to own cloud stability and incident response for a growing B2B SaaS platform. You will drive automation, observability, and tooling adoption while collaborating with cross-functional teams.

You will work primarily with AWS and on-premise VMs, implementing IaC with Terraform, Kubernetes, and monitoring stacks such as Prometheus, Grafana, and Datadog. On-site presence is required two days per week in Milton Keynes.

Qualifications

  • Previous experience as a Site Reliability Engineer or similar (cloud, infrastructure, DevOps or platform engineering) within a customer facing environment - e.g. SaaS or MSP.
  • Strong hands-on experience with both AWS, and on-premise virtual machines.
  • Experience with Infrastructure as Code/Terraform, Container orchestration (Kubernetes), and Monitoring and observability tooling (Prometheus, Grafana, Datadog, or Azure Monitor).
  • Ability to implement new processes, and tools, ensuring the wider development and support teams adopts new ways of working.
  • Ability to communicate across internal teams and external customers.
  • Skilled in networking across both cloud (Azure) and on premise environments.
  • Either Windows or Linux systems administration skills (Linux preferred).
  • Previous use of AI tools to enhance efficiency.

Responsibilities

  • Utilise various technologies (Terraform, Kubernetes ect) to manage provision, and configure Servers and networks, and automate application life cycles.
  • Regularly use Datadog and other observability tools for application performance monitoring.
  • Implement new ways of working, helping to shape how the organisation responds and recovers to incidents.
  • Take ownership of incident resolutions.
  • Actively drive down key reliability metrics (MTTR, incident frequency, on-call toil) by evaluating key incidents.
  • Work on an a on call rota, ensuring you are available to respond to incidents during this time.
  • Identify areas for automation and help implement changes that raise the bar for reliability.

Skills

Cloud & DevOps
Incident response
Cross-team communication
On-call rotation
Cloud networking
Linux/Windows admin
AI tooling usage

Tools

AWS
Terraform
Kubernetes
Prometheus
Grafana
Datadog
Azure Monitor

Job description

VIQU Ltd is seeking a Senior Site Reliability Engineer to own cloud stability and incident response for a growing B2B SaaS platform. You will drive automation, observability, and tooling adoption while collaborating with cross-functional teams.

You will work primarily with AWS and on-premise VMs, implementing IaC with Terraform, Kubernetes, and monitoring stacks such as Prometheus, Grafana, and Datadog. On-site presence is required two days per week in Milton Keynes.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AWS SRE: Platform Reliability & Automation
Senior AWS SRE: Platform Reliability & Automation

VIQU IT • Milton Keynes

On-site
GBP 68,000 - 83,000
On-site 2 days/week
Bonus
On-call allowance
+1
Senior Site Reliability Engineer — Reliability Lead
Senior Site Reliability Engineer — Reliability Lead

VIQU IT Recruitment • Milton Keynes

On-site
GBP 45,000 - 75,000
Senior SRE (AWS)
Senior SRE (AWS)

VIQU Ltd • Milton Keynes

On-site
GBP 60,000 - 75,000
Bonus
On-call allowance
On-site 2 days/week
Senior SRE (AWS)
Senior SRE (AWS)

VIQU IT • Milton Keynes

On-site
GBP 68,000 - 83,000
On-site 2 days/week
Bonus
On-call allowance
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VIQU IT Recruitment • Milton Keynes

On-site
GBP 45,000 - 75,000
Senior SRE - AWS, Kubernetes & Observability
Senior SRE - AWS, Kubernetes & Observability

ReVybe IT Recruitment Limited • Greater London

Hybrid
GBP 51,000 - 85,000
Bonus
Benefits
SRE Lead: CloudOps, IaC & 24/7 Reliability
SRE Lead: CloudOps, IaC & 24/7 Reliability

IQVIA Argentina • Greater London

Hybrid
GBP 90,000 - 125,000
Remote AWS SRE: Build Resilient Cloud Platforms
Remote AWS SRE: Build Resilient Cloud Platforms

SPECTRUM IT • England

On-site
GBP 70,000 - 110,000
Fully remote (UK)
24/7 shift pattern
Bonus & benefits
SRE Lead – CloudOps & Platform Reliability
SRE Lead – CloudOps & Platform Reliability

IQVIA Holdings Inc. • City of Westminster

On-site
GBP 110,000 - 150,000
Senior SRE & DevTools Engineer (CI/CD & Observability)
Senior SRE & DevTools Engineer (CI/CD & Observability)

Visa • Ham

Hybrid
GBP 70,000 - 110,000
Hybrid work model
Career development opportunities