ML Efficiency Engineer - Optimize Training & Serving

EngineersOfAI

United Kingdom

On-site

GBP 60,000 - 85,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

EngineersOfAI is seeking an experienced software engineer to join the ML Efficiency team. The role involves designing systems that enhance the efficiency of machine learning training and inference workloads. Candidates should possess a strong background in software engineering, particularly with Python and distributed systems.

Successful applicants will have the opportunity to impact performance improvements across the company's ML ecosystem, with a focus on optimizing resource utilization and developer productivity.

Qualifications

  • 5+ years of software engineering experience.
  • Experience building distributed systems at scale.
  • Experience with machine learning infrastructure and model serving platforms.

Responsibilities

  • Design and build systems that improve ML training and inference efficiency.
  • Develop tooling to help ML engineers debug and optimize model performance.
  • Partner with ML researchers to identify bottlenecks and drive performance improvements.
  • Optimize distributed training infrastructure and model serving architectures.

Skills

Python
Systems programming (Go, C++, Rust, or Java)
Performance engineering
Debugging and profiling

Education

BS, MS, or PhD in Computer Science or a related field

Tools

PyTorch Distributed
Ray
TensorFlow
Spark

Job description

EngineersOfAI is seeking an experienced software engineer to join the ML Efficiency team. The role involves designing systems that enhance the efficiency of machine learning training and inference workloads. Candidates should possess a strong background in software engineering, particularly with Python and distributed systems.

Successful applicants will have the opportunity to impact performance improvements across the company's ML ecosystem, with a focus on optimizing resource utilization and developer productivity.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Systems Performance Engineer
ML Systems Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Staff ML Performance Engineer — Edge Inference Optimizer
Staff ML Performance Engineer — Edge Inference Optimizer

Wayve • Greater London

Hybrid
GBP 70,000 - 90,000
Staff Engineer — ML Validation & Benchmarking
Staff Engineer — ML Validation & Benchmarking

EngineersOfAI • Greater London

On-site
GBP 60,000 - 80,000
Unlimited annual leave
Matched pension up to 5%
Phantom equity
+4
Production ML Engineer - Scalable AI Solutions
Production ML Engineer - Scalable AI Solutions

Faculty • Greater London

Hybrid
GBP 90,000 - 120,000
ML Platform Lead - LLM Training & Inference
ML Platform Lead - LLM Training & Inference

Scale AI • York and North Yorkshire

On-site
GBP 120,000 - 180,000
Health & Wellbeing
Career Growth stipend
Community events
+1
Staff Engineer — ML Validation & Benchmarking
Staff Engineer — ML Validation & Benchmarking

EngineersOfAI • Bristol

Hybrid
GBP 60,000 - 85,000
Unlimited annual leave
Up to 5% matched pension
Phantom equity
+5
Machine Learning Engineer
Machine Learning Engineer

Global One • Greater London

On-site
GBP 60,000 - 80,000
Build Infrastructure Engineer for ML Software Stack
Build Infrastructure Engineer for ML Software Stack

EngineersOfAI • Greater London

On-site
GBP 45,000 - 75,000
ML Engineer: Distributed Training & Low-Latency Inference
ML Engineer: Distributed Training & Low-Latency Inference

IMC Trading • Greater London

On-site
GBP 70,000 - 90,000
Senior AI & Automation Engineer - Build Scalable AI Solutions
Senior AI & Automation Engineer - Build Scalable AI Solutions

Optimizely • City of Westminster

On-site
GBP 70,000 - 100,000