Senior ML Engineer — Real-Time Distributed Training

IMC

Greater London

On-site

GBP 90,000 - 140,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

IMC is seeking a Machine Learning Engineer in London to help design and deploy large-scale ML models for trading applications. You will build distributed training pipelines, optimize low-latency inference, and collaborate with researchers and HPC specialists to accelerate experimentation and reduce costs.

You will work with Python, CUDA, and C++, and leverage PyTorch/TensorFlow/JAX, Horovod/NCCL, CuDNN, TensorRT in a high-performance environment.

Qualifications

  • 5+ years of ML experience with training or inference systems.
  • Hands-on experience with real-time, low-latency pipelines in high-performance envs.
  • Strong engineering skills: Python, CUDA, or C++.
  • Knowledge of PyTorch, TensorFlow, or JAX.
  • Proficiency in GPU programming for training/inference acceleration.

Responsibilities

  • Develop large-scale distributed training pipelines for datasets and models.
  • Build and optimize low-latency inference pipelines for real-time production.
  • Create libraries to improve ML framework performance.
  • Maximize training/inference performance using GPUs and acceleration libraries.
  • Collaborate with researchers to automate experiments and model retraining.
  • Partner with HPC specialists to optimize workflows and reduce costs.

Skills

Python
CUDA / C++
PyTorch
TensorFlow / JAX
GPU programming
Distributed training
Cloud platforms
Open-source contributions
Low-latency pipelines

Tools

Horovod
NCCL
TensorRT
CuDNN

Job description

IMC is seeking a Machine Learning Engineer in London to help design and deploy large-scale ML models for trading applications. You will build distributed training pipelines, optimize low-latency inference, and collaborate with researchers and HPC specialists to accelerate experimentation and reduce costs.

You will work with Python, CUDA, and C++, and leverage PyTorch/TensorFlow/JAX, Horovod/NCCL, CuDNN, TensorRT in a high-performance environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Engineer: Distributed Training & Low-Latency Inference
ML Engineer: Distributed Training & Low-Latency Inference

IMC Trading • Greater London

On-site
GBP 70,000 - 90,000
ML Engineer: Distributed Training & Low-Latency Inference
ML Engineer: Distributed Training & Low-Latency Inference

IMC • Greater London

On-site
GBP 75,000 - 115,000
ML Engineer: Scalable Training & Low-Latency Inference
ML Engineer: Scalable Training & Low-Latency Inference

IMC • Greater London

On-site
GBP 75,000 - 115,000
Machine Learning Engineer
Machine Learning Engineer

IMC Trading • Greater London

On-site
GBP 70,000 - 90,000
Machine Learning Engineer
Machine Learning Engineer

IMC • Greater London

On-site
GBP 90,000 - 140,000
ML Infrastructure Engineer - Distributed Systems & HPC
ML Infrastructure Engineer - Distributed Systems & HPC

Stanford Black Limited • Greater London

On-site
GBP 100,000 - 150,000
Competitive compensation
Bonus structure
Autonomy from day one
+1
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
Applied ML Engineer: Real-World Ops & Deployment
Applied ML Engineer: Real-World Ops & Deployment

Platform Recruitment • Greater London

On-site
GBP 42,000 - 70,000
Senior ML Engineer – Hybrid London (CUDA, PyTorch)
Senior ML Engineer – Hybrid London (CUDA, PyTorch)

Hexwired Recruitment Limited • Greater London

Hybrid
GBP 83,000 - 110,000
ML Performance Engineer: Scale GPU/CPU ML Workloads
ML Performance Engineer: Scale GPU/CPU ML Workloads

G-Research • Greater London

Hybrid
GBP 90,000 - 150,000
Competitive pay
Lunch provided
Annual leave 35d
+5