Hybrid Site Reliability Engineer — Scale & Automate Cloud

Capital on Tap

Greater London

Hybrid

GBP 70,000 - 120,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Private Healthcare including dental &_
Worldwide travel insurance
Anniversary Rewards (£250, £500, £750)
Salary Sacrifice Pension Scheme up to
28 days holiday
Annual Learning and Wellbeing Budget
Enhanced Parental Leave
Cycle to Work Scheme
Season Ticket Loan
6 free therapy sessions per year
Dog Friendly Offices
Free drinks and snacks

Job summary

Capital On Tap is hiring a Site Reliability Engineer to join a hybrid embedded SRE model. You will design, build, monitor and scale platforms, collaborating with XP to own architecture and paved paths for product teams.

You will manage cloud resources, implement IaC, and enhance CI/CD pipelines while improving system reliability and performance across the stack. The role emphasizes proactive incident response and strong stakeholder collaboration.

Qualifications

  • Experience managing public cloud environments and related services.
  • Proficient in writing, managing and optimising infrastructure using IaC tools like Terraform.
  • Experience with CI/CD tools, pipelines, templates and troubleshooting.
  • Proficient with containerisation (Kubernetes and Docker).
  • Experience building/deploying frontend applications with CodeMagic.
  • Knowledge or experience with Google Firebase.
  • Experience with cloud monitoring solutions.
  • Proficiency in at least one scripting language (Python, PowerShell, Go).
  • Strong communication and collaboration skills; software development background.

Responsibilities

  • Manage and automate resources in Azure, Datadog, NGINX & Cloudflare.
  • Develop, deploy and monitor Kubernetes and Serverless resources.
  • Build, manage and evolve IaC using Terraform, Helm and Go CRDs.
  • Improve systems, processes and technologies; consult stakeholders on performance.
  • Design and maintain monitoring and alerting strategies using Datadog.
  • Contribute to new application architecture and design processes.
  • Design solutions to reduce toil and automate tasks to boost productivity.
  • Create SLIs/SLOs and increase application visibility.
  • Align with Product on SLAs and core service objectives.
  • Collaborate with core teams to build reusable, automated solutions.
  • Optimise CI/CD pipelines and developer workflows with Azure DevOps, Github, Octopus Deploy, MirrorD, Flux.
  • Lead incident response and participate in post-mortems to protect customer experience.

Skills

Public cloud
Kubernetes
Docker
CI/CD
Infrastructure as Code
Frontend build tooling (CodeMagic)
Firebase
Cloud monitoring
Scripting (Python/PowerShell/Go)
Software development background
Communication

Tools

Terraform
Helm
Go CRDs
Azure DevOps
Github
Octopus Deploy
MirrorD
Flux

Job description

Capital On Tap is hiring a Site Reliability Engineer to join a hybrid embedded SRE model. You will design, build, monitor and scale platforms, collaborating with XP to own architecture and paved paths for product teams.

You will manage cloud resources, implement IaC, and enhance CI/CD pipelines while improving system reliability and performance across the stack. The role emphasizes proactive incident response and strong stakeholder collaboration.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Azure, Kubernetes & IaC Expert
Site Reliability Engineer - Azure, Kubernetes & IaC Expert

Capitalontap • Greater London

Hybrid
GBP 55,000 - 85,000
Private Healthcare
Worldwide travel insurance
Anniversary Rewards
+9
Hybrid SRE: Build Reliable, Scalable Systems
Hybrid SRE: Build Reliable, Scalable Systems

bet365 • Manchester

Hybrid
GBP 70,000 - 110,000
Hybrid work from home policy
Senior Cloud Reliability Engineer — Hybrid & Automation Focus
Senior Cloud Reliability Engineer — Hybrid & Automation Focus

LLOYDS BANKING GROUP • Leeds

Hybrid
GBP 90,000 - 120,000
Annual bonus
Share options
30 days holiday
+1
Senior SRE: Reliability & Scale for Banking (Hybrid)
Senior SRE: Reliability & Scale for Banking (Hybrid)

GCS • Glasgow

Hybrid
GBP 75,000 - 110,000
Hybrid Site Reliability Engineer — Healthcare Infra & Automation
Hybrid Site Reliability Engineer — Healthcare Infra & Automation

Worky • Greater London

Hybrid
GBP 85,000 - 107,000
Medical, Vision, and Dental Insurance
401k Retirement Plan
3.5 Weeks Paid Time Off plus 7 company
+2
Senior Site Reliability Engineer & DevTools Specialist
Senior Site Reliability Engineer & DevTools Specialist

Visa • Basingstoke

Hybrid
GBP 70,000 - 120,000
Senior Cloud SRE: Kubernetes, GCP & Reliable Automation
Senior Cloud SRE: Kubernetes, GCP & Reliable Automation

Jobs Paloaltonetworks • United Kingdom

On-site
GBP 90,000 - 120,000
Senior Site Reliability Engineer — Hybrid, Global Impact
Senior Site Reliability Engineer — Hybrid, Global Impact

twentysix • United Kingdom

Hybrid
GBP 51,000 - 77,000
Hybrid work options
25 vacation days
Transportation subsidies
+6
Senior Site Reliability Engineer - Scale & Automation
Senior Site Reliability Engineer - Scale & Automation

Google • Greater London

On-site
GBP 110,000 - 170,000
AI-Driven Site Reliability Engineer - Scalable Cloud Infra
AI-Driven Site Reliability Engineer - Scalable Cloud Infra

Cisco • Greater London

On-site
GBP 90,000 - 130,000