Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Worky

Birmingham

On-site

GBP 150,000 - 210,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services.

The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context.

Qualifications

  • Hands-on experience building and operating reliable distributed systems.
  • Experience with IaC, container orchestration, and cloud-native architectures.
  • Strong debugging, incident response, and performance tuning skills.
  • Bachelor’s degree in CS or related field; 7+ years in SRE/DevOps roles.

Responsibilities

  • Partner with engineering leadership to define SLOs, SLIs and error budgets.
  • Architect highly available, fault-tolerant systems and review patterns like circuit breakers and rate limits.
  • Reduce toil by building automation, tooling, and self-service capabilities.
  • Improve production readiness through load testing, capacity forecasting and chaos engineering.
  • Lead complex multi-system incidents and drive blameless post-mortems for root-cause analysis.
  • Design sustainable on-call models with clear escalation and balanced pager responsibilities.

Skills

Java/Python/Node.js
IaC (Terraform/Ansible/CloudFormation)
Docker
Kubernetes
Cloud platforms (AWS/GCP/Azure)
Observability (Prometheus/Grafana/Open
Linux/SDLC & Testing
Networking & Load Balancing

Education

Bachelor’s degree in Computer Science or related field

Tools

Terraform
Ansible
CloudFormation
Docker
Kubernetes
Prometheus
Grafana
OpenTelemetry
ELK/Elastic
CloudWatch

Job description

WHAT WE DO

Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help engineer highly reliable, observable, and resilient platforms that support critical business services at scale. You will collaborate with multiple engineering teams to continually improve our production system architecture, facilitate fast delivery of new services, and reduce downtime.

This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.

Key Responsibilities
  • Partner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.
  • Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.
  • Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.

  • Improve production readiness through load testing, performance tuning, capacity forecasting, chaos engineering and reliability reviews.

  • Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.

  • Promote sustainable operations by helping design healthy on-call models, clear escalation paths, and balanced pager responsibilities.

WHAT WE ARE LOOKING FOR
Core Technical Skills
  • Strong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.
  • Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.

  • Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.

  • Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.

  • Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)
  • Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.

  • Knowledge of networking protocols(VPC) , load balancing strategies in a distributed systems environment , database query performance tuning and identify cause for lagging

Core Competencies & Soft Skills
  • Ability to analyze complex, distributed systems holistically and understand how individual components interact under load.

  • Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.

  • Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.
  • Highly motivated, pro-active and capable of multi-tasking under pressure in a fast-paced environment without compromising quality.

  • Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.

  • Interest in financial markets and technology.

Preferred Qualifications
  • Bachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.

  • 7 to 10 years of experience

ABOUT GOLDMAN SACHS

The Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded in 1869, the firm is headquartered in New York and maintains offices in all major financial centers around the world.

At Goldman Sachs, we commit our people, capital and ideas to help our clients, shareholders and the communities we serve to grow. Founded in 1869, we are a leading global investment banking, securities and investment management firm. Headquartered in New York, we maintain offices around the world.

We believe who you are makes you better at what you do. We're committed to fostering and advancing diversity and inclusion in our own workplace and beyond by ensuring every individual within our firm has a number of opportunities to grow professionally and personally, from our training and development opportunities and firmwide networks to benefits, wellness and personal finance offerings and mindfulness programs. Learn more about our culture, benefits, and people at GS.com/careers.

We’re committed to finding reasonable accommodations for candidates with special needs or disabilities during our recruiting process. Learn more: https://www.goldmansachs.com/careers/footer/disability-statement.html

© The Goldman Sachs Group, Inc., 2026. All rights reserved.

Goldman Sachs is an equal opportunity employer and does not discriminate on the basis of race, color, religion, sex, national origin, age, veterans status, disability, or any other characteristic protected by applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

The Core Engineering - Site Reliability Engineering - Associate - Birmingham
The Core Engineering - Site Reliability Engineering - Associate - Birmingham

WeAreTechWomen • Birmingham

On-site
GBP 65,000 - 90,000
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Goldman Sachs Bank AG • Birmingham

Hybrid
GBP 120,000 - 180,000
Healthcare & Medical Insurance
Financial Wellness & Retirement
On-site health centers
The Core Engineering - Software Engineer - Associate - Birmingham
The Core Engineering - Software Engineer - Associate - Birmingham

Goldman Sachs • West Midlands

On-site
GBP 70,000 - 100,000
The Core Engineering - Software Engineer - Associate - Birmingham
The Core Engineering - Software Engineer - Associate - Birmingham

WeAreTechWomen • Birmingham

On-site
GBP 70,000 - 110,000
IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]
IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]

Goldman Sachs Bank AG • Greater London

On-site
GBP 80,000 - 110,000
Asset & Wealth Management - Software Engineering Lead - Vice President - Birmingham
Asset & Wealth Management - Software Engineering Lead - Vice President - Birmingham

Goldman Sachs • Birmingham

On-site
GBP 90,000 - 130,000
Asset & Wealth Management - Software Engineering Lead - Vice President - Birmingham
Asset & Wealth Management - Software Engineering Lead - Vice President - Birmingham

Goldman Sachs • West Midlands

On-site
GBP 70,000 - 90,000
VP, Site Reliability Engineering: Scale & Resilience
VP, Site Reliability Engineering: Scale & Resilience

Worky • Birmingham

On-site
GBP 150,000 - 210,000
Engineering - Regional COO - Vice President – London London · United Kingdom · Vice President
Engineering - Regional COO - Vice President – London London · United Kingdom · Vice President

Goldman Sachs Bank AG • Greater London

On-site
GBP 80,000 - 100,000
Medical, dental, and disability insurance
Retirement planning support
On-site fitness centers
+1
Software Engineer | Associate | London
Software Engineer | Associate | London

Goldman Sachs • Greater London

On-site
GBP 70,000 - 95,000