Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey

Greater London

On-site

GBP 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Odyssey in Greater London seeks an experienced software engineer specializing in machine learning performance optimization. You will optimize models for real-time users, design distributed training strategies, and work with elite ML researchers. Candidates should have at least 8 years of software engineering experience, deep insights into machine learning architectures, and proficiency in PyTorch and NVIDIA optimization. This position offers autonomy in technical decisions and a chance to work with cutting-edge technology.

Qualifications

  • 8+ years of software engineering experience with significant work in ML performance.
  • Deep insight into modern machine learning architectures.
  • Track record of owning projects end to end.
  • Proficiency with PyTorch, Triton, and NVIDIA GPU ecosystems.

Responsibilities

  • Optimize models for real-time use by hundreds of thousands of users.
  • Design distributed training strategies for GPU clusters.
  • Develop tools to identify performance bottlenecks.
  • Pioneer innovative approaches to enhance performance metrics.

Skills

Software engineering
Machine learning performance optimization
Distributed training
PyTorch
NVIDIA GPU optimization

Tools

Triton
TF/JAX

Job description

Odyssey in Greater London seeks an experienced software engineer specializing in machine learning performance optimization. You will optimize models for real-time users, design distributed training strategies, and work with elite ML researchers. Candidates should have at least 8 years of software engineering experience, deep insights into machine learning architectures, and proficiency in PyTorch and NVIDIA optimization. This position offers autonomy in technical decisions and a chance to work with cutting-edge technology.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
Realtime ML Systems Engineer - High-Performance Inference
Realtime ML Systems Engineer - High-Performance Inference

Inworld AI • United Kingdom

On-site
GBP 140,000 - 200,000
Senior ML Engineer — Real-Time Distributed Training
Senior ML Engineer — Real-Time Distributed Training

IMC • Greater London

On-site
GBP 90,000 - 140,000
Staff Engineer, ML Infrastructure & Compute Platform
Staff Engineer, ML Infrastructure & Compute Platform

Odyssey • Greater London

On-site
GBP 60,000 - 80,000
ML Performance Engineer: Scale GPU/CPU ML Workloads
ML Performance Engineer: Scale GPU/CPU ML Workloads

G-Research • Greater London

Hybrid
GBP 90,000 - 150,000
Competitive pay
Lunch provided
Annual leave 35d
+5
Senior ML Engineer - Real-Time Inference at Scale
Senior ML Engineer - Real-Time Inference at Scale

careers.bitkraft.vc - Jobboard • United Kingdom

On-site
GBP 140,000 - 200,000
Research Engineer: Scalable Production ML Systems
Research Engineer: Scalable Production ML Systems

Anthropic • Greater London

On-site
GBP 260,000 - 630,000
Remote Performance Engineer: ML Training & Kernels
Remote Performance Engineer: ML Training & Kernels

Cohere • Greater London

On-site
GBP 75,000 - 95,000
Co-working benefit
Daily lunch program
Regular community and social events
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4
Remote Staff ML Efficiency Engineer — Scale & Optimize
Remote Staff ML Efficiency Engineer — Scale & Optimize

Reddit, Inc. • Greater London

On-site
GBP 110,000 - 160,000
Global Benefits
Family Planning
Mental Health Support
+4