Senior ML Systems Engineer: Large-Scale Performance

Stanford Black Limited

England

On-site

GBP 70,000 - 150,000

Full time

9 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Stanford Black Limited partners with a highly quantitative research organisation to build large-scale machine learning systems in a performance-critical environment. The role sits at the intersection of ML, distributed systems, and high-performance computing, focusing on scaling modern ML workloads and improving training and inference efficiency for large models.

Responsibilities include designing and optimising training and inference systems, improving throughput, latency, memory efficiency,

Qualifications

  • 6+ years’ experience in Machine Learning Engineering, Research Engineering, ML Infrastructure, Distributed Systems, or Performance Engineering.
  • Strong Python and/or C++ development experience.
  • Deep understanding of modern ML frameworks including PyTorch, JAX, or TensorFlow.
  • Experience training, deploying, or optimising large-scale machine learning models.
  • Strong understanding of parallel computing, distributed systems, and performance optimisation.
  • Degree (or equivalent experience) in Computer Science, Mathematics, Physics, Engineering, or a related quantitative discipline.

Responsibilities

  • Design and optimise large-scale training and inference systems.
  • Improve throughput, latency, memory efficiency, and GPU utilisation across distributed workloads.
  • Partner with researchers to translate new ML ideas into scalable production systems.
  • Build infrastructure and tooling that accelerates experimentation, model development, and deployment.
  • Drive technical direction across performance-critical ML systems and compute infrastructure.
  • Solve challenging problems spanning software, hardware, compilers, and distributed computing.

Skills

Python
C++
Machine Learning
Distributed Systems
Performance Engineering

Education

Degree in CS/Math/Physics/Engineering

Tools

PyTorch
JAX
TensorFlow
DeepSpeed/Megatron

Job description

Stanford Black Limited partners with a highly quantitative research organisation to build large-scale machine learning systems in a performance-critical environment. The role sits at the intersection of ML, distributed systems, and high-performance computing, focusing on scaling modern ML workloads and improving training and inference efficiency for large models.

Responsibilities include designing and optimising training and inference systems, improving throughput, latency, memory efficiency,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

High-Throughput ML Systems Performance Engineer
High-Throughput ML Systems Performance Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Fertility benefits
Parental leave 22 weeks
+12
ML Systems Performance Engineer
ML Systems Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Performance Engineer
Performance Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Fertility benefits
Parental leave 22 weeks
+12
Real-Time ML Systems Performance Engineer
Real-Time ML Systems Performance Engineer

Janestreet • Greater London

On-site
GBP 70,000 - 90,000
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4
Senior ML Systems Engineer - Scale & Performance
Senior ML Systems Engineer - Scale & Performance

Arcus Search • Greater London

On-site
GBP 80,000 - 120,000
Realtime ML Systems Engineer - High-Performance Inference
Realtime ML Systems Engineer - High-Performance Inference

Inworld AI • United Kingdom

On-site
GBP 140,000 - 200,000
Senior ML Systems Engineer - Simulations
Senior ML Systems Engineer - Simulations

Oriole • Greater London

On-site
GBP 90,000 - 150,000
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
LLM Performance Engineer — Scale, Optimize & Deploy
LLM Performance Engineer — Scale, Optimize & Deploy

Isomorphic Labs • Greater London

Hybrid
GBP 120,000 - 160,000