Staff AI Systems Engineer - Pre-Training Infra

Reflection

Greater London

On-site

GBP 70,000 - 100,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Top-tier compensation
Comprehensive health insurance
Fully paid parental leave
Paid time off
Daily lunch and dinner provided

Job summary

Reflection in Greater London is seeking a talented engineer to build and scale distributed training systems for cutting-edge machine learning models. The role involves optimizing GPU performance and collaborating with research teams to create efficient training infrastructures. Ideal candidates should have experience with modern distributed frameworks and large dataset pipelines. The position offers a strong compensation package, comprehensive health benefits, and a supportive work environment to do impactful work.

Qualifications

  • Experience in building or operating large ML models.
  • Strong familiarity with distributed training frameworks.
  • Proven skills in optimizing ML workloads and GPUs.

Responsibilities

  • Build and scale distributed training systems for ML models.
  • Collaborate with researchers on large training runs.
  • Optimize training throughput and efficiency across GPUs.

Skills

Experience building or operating distributed training systems for large machine learning models
Strong experience with modern distributed training frameworks
Familiarity with large-scale model parallelism strategies
Experience optimizing training throughput and GPU utilization
Familiarity with GPU communication libraries
Experience working closely with ML researchers
Strong debugging skills across GPU compute
Experience working with large datasets

Job description

Reflection in Greater London is seeking a talented engineer to build and scale distributed training systems for cutting-edge machine learning models. The role involves optimizing GPU performance and collaborating with research teams to create efficient training infrastructures. Ideal candidates should have experience with modern distributed frameworks and large dataset pipelines. The position offers a strong compensation package, comprehensive health benefits, and a supportive work environment to do impactful work.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Software Engineer - ML Training Infrastructure
Research Software Engineer - ML Training Infrastructure

Reflection • Greater London

On-site
GBP 70,000 - 100,000
Senior ML Infra Architect for Large-Scale AI Simulations
Senior ML Infra Architect for Large-Scale AI Simulations

Hamilton Barnes Associates Limited • United Kingdom

Hybrid
GBP 90,000 - 130,000
Significant stock option packages
Remote-first working setup
Fully paid travel and accommodation
+1
Platform Systems Engineer: Scalable AI Training & Telemetry
Platform Systems Engineer: Scalable AI Training & Telemetry

OpenAI • City Of London

On-site
GBP 70,000 - 90,000
Staff AI Engineer: Post-Training, RL & Model Alignment
Staff AI Engineer: Post-Training, RL & Model Alignment

Reflection AI • Greater London

On-site
GBP 75,000 - 110,000
Top-tier compensation
Comprehensive medical, dental, vision, life, and disability insurance
Fully paid parental leave
+2
Platform Systems Engineer: Scalable AI Infra & Observability
Platform Systems Engineer: Scalable AI Infra & Observability

OpenAI • Greater London

On-site
GBP 60,000 - 80,000
Lead ML Infra Engineer for Scalable Physics AI
Lead ML Infra Engineer for Scalable Physics AI

PhysicsX • City Of London

On-site
GBP 80,000 - 120,000
Senior AI Infrastructure Engineer — Scale Training & Inference
Senior AI Infrastructure Engineer — Scale Training & Inference

Fin • Greater London

Hybrid
GBP 90,000 - 150,000
Systems Research Engineer (AI Infrastructure & Distributed Systems
Systems Research Engineer (AI Infrastructure & Distributed Systems

European Tech Recruit • City of Edinburgh

On-site
GBP 60,000 - 80,000
Senior GPU Systems Engineer: Large-Scale Inference & RL
Senior GPU Systems Engineer: Large-Scale Inference & RL

Reflection • Greater London

On-site
GBP 70,000 - 100,000
Research Engineer – Large-Scale Pretraining & Systems
Research Engineer – Large-Scale Pretraining & Systems

Anthropic • Greater London

On-site
GBP 250,000 - 435,000
Equity benefits
Visa sponsorship