Real-Time ML Systems Performance Engineer

Janestreet

Greater London

On-site

GBP 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Janestreet is seeking an engineer with a focus on low-level systems programming and optimisation to join our ML team in Greater London. The role involves enhancing the performance of machine learning models, ensuring efficient large-scale training and real-time inference.

Candidates should have strong knowledge in GPU technologies and an understanding of modern machine learning techniques. Debugging, optimisation skills, and fluency in English are essential for this position.

Qualifications

  • Experience in low-level systems programming and optimisation.
  • The experience and systems knowledge to debug training runs end to end.
  • Intuitive understanding of CUDA memory hierarchy and GPU performance.

Responsibilities

  • Optimise the performance of ML models for training and inference.
  • Ensure efficient large-scale training in the trading environment.
  • Improve systems-level performance and throughput analytics.

Skills

Understanding of modern ML techniques and toolsets
Low-level GPU knowledge
Debugging and optimisation experience
Library knowledge of Triton, CUTLASS, CUB
Intuition about latency and throughput in CUDA
Background in Infiniband, RoCE, and related networking
Fluency in English

Tools

CUDA GDB
NSight Systems
cuDNN
cuBLAS

Job description

Janestreet is seeking an engineer with a focus on low-level systems programming and optimisation to join our ML team in Greater London. The role involves enhancing the performance of machine learning models, ensuring efficient large-scale training and real-time inference.

Candidates should have strong knowledge in GPU technologies and an understanding of modern machine learning techniques. Debugging, optimisation skills, and fluency in English are essential for this position.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Performance Engineer: GPU & Systems Optimisation
ML Performance Engineer: GPU & Systems Optimisation

Trading Interview • Greater London

Hybrid
GBP 120,000 - 180,000
ML Systems Performance Engineer
ML Systems Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Machine Learning Performance Engineer
Machine Learning Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Machine Learning Performance Engineer
Machine Learning Performance Engineer

Trading Interview • Greater London

Hybrid
GBP 120,000 - 180,000
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
Realtime ML Systems Engineer - High-Performance Inference
Realtime ML Systems Engineer - High-Performance Inference

Inworld AI • United Kingdom

On-site
GBP 140,000 - 200,000
ML Performance Engineer: Scale GPU/CPU ML Workloads
ML Performance Engineer: Scale GPU/CPU ML Workloads

G-Research • Greater London

Hybrid
GBP 90,000 - 150,000
Competitive pay
Lunch provided
Annual leave 35d
+5
Senior ML Engineer — Real-Time Distributed Training
Senior ML Engineer — Real-Time Distributed Training

IMC • Greater London

On-site
GBP 90,000 - 140,000
Remote Performance Engineer: ML Training & Kernels
Remote Performance Engineer: ML Training & Kernels

Cohere • Greater London

On-site
GBP 75,000 - 95,000
Co-working benefit
Daily lunch program
Regular community and social events
Senior ML Engineer - Low-Latency Production Pipelines
Senior ML Engineer - Low-Latency Production Pipelines

Longshot Systems • Greater London

Hybrid
GBP 90,000 - 140,000
Participation in uncapped bonus
10% matched pension
Private healthcare
+2