Senior GPU Systems Engineer - Clusters & Platform Infra

Radley James

Greater London

On-site

GBP 20,000 - 40,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Radley James is seeking a Senior Systems Engineer in London to lead cluster management and platform engineering for large GPU deployments. The role emphasizes reliability, monitoring, and scalable infrastructure to support growing AI workloads.

The ideal candidate has extensive experience with multi-thousand GPU environments, demonstrated project leadership, and a track record of delivering mission-critical infrastructure in fast-paced teams.

Qualifications

  • 6+ years experience in a high performance field such as AI, big tech, or quantitative trading.
  • Experience working on clusters of 1000 GPUs or higher.
  • Experience driving key projects in your team or business.

Responsibilities

  • Focus on cluster management and platform engineering for large GPU deployments.
  • Improve monitoring and reliability of GPU-based infrastructures.
  • Drive infrastructure initiatives for next-generation GPU deployments.

Skills

6 years experience in high performance
1000+ GPU clusters
Project leadership

Job description

Radley James is seeking a Senior Systems Engineer in London to lead cluster management and platform engineering for large GPU deployments. The role emphasizes reliability, monitoring, and scalable infrastructure to support growing AI workloads.

The ideal candidate has extensive experience with multi-thousand GPU environments, demonstrated project leadership, and a track record of delivering mission-critical infrastructure in fast-paced teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

HPC & AI Platform Engineer – GPU Clusters
HPC & AI Platform Engineer – GPU Clusters

Era4 • United Kingdom

Hybrid
GBP 95,000 - 130,000
Site Reliability Engineer, GPUs in AI
Site Reliability Engineer, GPUs in AI

Radley James • Greater London

On-site
GBP 20,000 - 40,000
Lead GPU Infrastructure Architect for Scalable AI Clusters
Lead GPU Infrastructure Architect for Scalable AI Clusters

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 110,000 - 150,000
GPU Infrastructure Lead - Systems Integrator
GPU Infrastructure Lead - Systems Integrator

Hamilton Barnes Associates Limited • Greater London

On-site
GBP 140,000 - 170,000
Full Benefits
Technical Lead, HPC & AI Infrastructure
Technical Lead, HPC & AI Infrastructure

Era4 • England

On-site
GBP 90,000 - 150,000
Platform Engineer — GPU HPC & Bare-Metal Clusters
Platform Engineer — GPU HPC & Bare-Metal Clusters

CATCHES • United Kingdom

Remote
GBP 50,000 - 70,000
Senior GPU & AI Infra Architect — Remote, 4-Day Week
Senior GPU & AI Infra Architect — Remote, 4-Day Week

Civo Ltd • United Kingdom

Hybrid
GBP 110,000 - 170,000
4-day week
Uncapped holidays
Remote work environment
Senior Systems Engineer: Performance & Reliability (Linux)
Senior Systems Engineer: Performance & Reliability (Linux)

Eworker • Bristol

Hybrid
GBP 70,000 - 110,000
Flexible working
Private medical insurance
Dental plan
+4
Senior GPU HPC Engineer: InfiniBand & KVM Optimization
Senior GPU HPC Engineer: InfiniBand & KVM Optimization

Nebius • Greater London

On-site
GBP 90,000 - 130,000
Competitive compensation
Career growth
Flexibility and ownership
+3
Senior Systems Engineer: Drive Performance & Reliability
Senior Systems Engineer: Drive Performance & Reliability

Eworker • United Kingdom

Hybrid
GBP 110,000 - 170,000
Flexible working
Generous annual leave
Private medical insurance
+7