VP, Site Reliability Engineering: Scale & Resilience

Worky

Birmingham

On-site

GBP 150,000 - 210,000

Full time

6 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services.

The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context.

Qualifications

  • Hands-on experience building and operating reliable distributed systems.
  • Experience with IaC, container orchestration, and cloud-native architectures.
  • Strong debugging, incident response, and performance tuning skills.
  • Bachelor’s degree in CS or related field; 7+ years in SRE/DevOps roles.

Responsibilities

  • Partner with engineering leadership to define SLOs, SLIs and error budgets.
  • Architect highly available, fault-tolerant systems and review patterns like circuit breakers and rate limits.
  • Reduce toil by building automation, tooling, and self-service capabilities.
  • Improve production readiness through load testing, capacity forecasting and chaos engineering.
  • Lead complex multi-system incidents and drive blameless post-mortems for root-cause analysis.
  • Design sustainable on-call models with clear escalation and balanced pager responsibilities.

Skills

Java/Python/Node.js
IaC (Terraform/Ansible/CloudFormation)
Docker
Kubernetes
Cloud platforms (AWS/GCP/Azure)
Observability (Prometheus/Grafana/Open
Linux/SDLC & Testing
Networking & Load Balancing

Education

Bachelor’s degree in Computer Science or related field

Tools

Terraform
Ansible
CloudFormation
Docker
Kubernetes
Prometheus
Grafana
OpenTelemetry
ELK/Elastic
CloudWatch

Job description

Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services.

The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

VP, SRE: Architect Resilient, Scalable Systems
VP, SRE: Architect Resilient, Scalable Systems

Goldman Sachs Bank AG • Birmingham

Hybrid
GBP 120,000 - 180,000
Healthcare & Medical Insurance
Financial Wellness & Retirement
On-site health centers
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Worky • Birmingham

On-site
GBP 150,000 - 210,000
Associate Site Reliability Engineer — Automate and Scale
Associate Site Reliability Engineer — Automate and Scale

Goldman Sachs • West Midlands

On-site
GBP 70,000 - 100,000
SRE Engineer: Core Systems & Automation
SRE Engineer: Core Systems & Automation

WeAreTechWomen • Birmingham

On-site
GBP 70,000 - 110,000
IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]
IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]

Goldman Sachs Bank AG • Greater London

On-site
GBP 80,000 - 110,000
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Goldman Sachs Bank AG • Birmingham

Hybrid
GBP 120,000 - 180,000
Healthcare & Medical Insurance
Financial Wellness & Retirement
On-site health centers
Lead Site Reliability Engineer — Incident & Resilience
Lead Site Reliability Engineer — Incident & Resilience

J.P. MORGAN • Greater London

On-site
GBP 140,000 - 180,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

JPMorgan Chase & Co. • City of Westminster

On-site
GBP 110,000 - 150,000
Senior SRE: AWS Platform Reliability Architect
Senior SRE: AWS Platform Reliability Architect

JPMorganChase • Glasgow

On-site
GBP 90,000 - 130,000
Senior Site Reliability & DevOps Lead
Senior Site Reliability & DevOps Lead

JPMorgan Chase & Co. • Auchentibber

On-site
GBP 90,000 - 130,000