SRE Engineer

Selby Jennings

Greater London

On-site

GBP 90,000 - 130,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Selby Jennings is hired by a world-renowned hedge fund to recruit an ambitious SRE who will drive reliability, scalability, and performance of a core platform. You will own projects from inception to production, with autonomy and a strong focus on IaC, automation, and observability.

The role balances critical operational support with new infrastructure initiatives, emphasizing proactive problem solving and strong engineering ownership within a high-velocity environment.

Qualifications

  • 4+ years of hands-on experience with Google Cloud Platform (GCP) or comparable cloud infrastructure.
  • Expert-level proficiency in managing production Kubernetes environments.
  • Deep expertise in Terraform for cloud and Kubernetes resources.
  • Strong experience with Helm for packaging and deploying applications on Kubernetes.
  • Proficiency in Python (and Bash) for automation and tooling.

Responsibilities

  • Design, build, and maintain scalable, reliable infrastructure on GCP using Kubernetes.
  • Champion automation across the SDLC with Infrastructure-as-Code (IaC) practices.
  • Own and evolve declarative infrastructure with Terraform and Helm.
  • Implement monitoring, alerting, and logging for clear system visibility.
  • Define and enforce SLOs/SLIs; participate in on-call rotation and post-incident reviews.

Skills

GCP experience
Kubernetes
Python
Bash
CI/CD
Observability
Linux
Networking
SRE

Tools

Terraform
Helm
Kubernetes
GCP

Job description

Our client, a world-renowned hedge fund, is looking for an SRE to join their agile, high-performing engineering team.This role offers a unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership. You will take full ownership of projects from inception through to production operations, including documentation and knowledge transfer.

The successful candidate will divide their time between providing expert operational support for our critical systems and leading exciting new infrastructure projects. Our mindset is to find the right person, not simply someone whose experience perfectly matches our current technology stack. The ability to anticipate problems before they arise and develop solutions to prevent them is key. If you enjoy a challenging environment, implementing Infrastructure-as-Code principles, and seeing the direct impact of your work, this is the place for you.

What You Will Do:
  • Design & Build: Architect, deploy, and maintain highly scalable and reliable infrastructure on Google Cloud Platform (GCP) using Kubernetes and Infrastructure-as-Code tools.
  • Automation: Champion automation across the entire software development lifecycle (SDLC), utilizing IaC, Python, and Bash to reduce manual effort and improve operational efficiency.
  • Infrastructure-as-Code (IaC): Own and evolve our declarative infrastructure using Terraform for cloud resources and Helm for Kubernetes application deployments.
  • Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue detection.
  • Reliability & Performance: Define, measure, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Participate in the on-call rotation, where applicable, and lead post-incident reviews to drive continuous improvement.
Required Experience & Skills
Core Technical Stack
  • Cloud Platform: 4+ years of hands-on experience with Google Cloud Platform (GCP) or a comparable cloud infrastructure provider.
  • Container Orchestration: Expert-level proficiency in managing, scaling, and troubleshooting production Kubernetes environments.
  • Infrastructure-as-Code: Deep expertise in Terraform for managing cloud and Kubernetes resources.
  • Deployment: Strong experience with Helm for packaging and deploying applications on Kubernetes.
  • Scripting/Programming: Proficiency in at least one major programming language, preferably Python, for automation and tool development.
Tooling & Concepts
  • CI/CD: Experience building, maintaining, and improving modern CI/CD pipelines.
  • Observability: Practical experience implementing and managing monitoring, alerting, and logging solutions.
  • Networking: Solid understanding of TCP/IP, load balancing, DNS, and cloud-native networking within Kubernetes environments.
  • Operating Systems: Strong command-line skills and hands-on experience with Linux systems.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Platform Reliability Engineer — IaC, Kubernetes & GCP
Platform Reliability Engineer — IaC, Kubernetes & GCP

Selby Jennings • Greater London

On-site
GBP 90,000 - 130,000
Senior Platform Engineer
Senior Platform Engineer

Selby Jennings • Greater London

On-site
GBP 90,000 - 130,000
Senior Platform DevOps Engineer - CI/CD, Kubernetes, Scale
Senior Platform DevOps Engineer - CI/CD, Kubernetes, Scale

Selby Jennings • Greater London

On-site
GBP 90,000 - 120,000
Senior DevOps Engineer
Senior DevOps Engineer

Selby Jennings • Greater London

On-site
GBP 90,000 - 120,000
SRE Engineer
SRE Engineer

Savant Recruitment • Greater London

On-site
GBP 60,000 - 80,000
Devops SRE
Devops SRE

Test Triangle • Greater London

On-site
GBP 70,000 - 90,000
SRE / DevOps Engineers
SRE / DevOps Engineers

HCLTech • Greater London

On-site
GBP 60,000 - 80,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Selby Jennings • Greater London

On-site
GBP 70,000 - 90,000
Senior SRE
Senior SRE

Pulse Recruit • Greater London

On-site
GBP 65,000 - 85,000
Senior Site Reliability Engineer (Platform Reliability, Resilience)
Senior Site Reliability Engineer (Platform Reliability, Resilience)

Elastic • Greater London

Hybrid
GBP 90,000 - 130,000
Health coverage for you and family
Flexible location & schedule
Generous vacation days
+3