Senior ML Performance Engineer - Real-Time Inference & Scale
Odyssey
Greater London
On-site
GBP 70,000 - 90,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
Odyssey in Greater London seeks an experienced software engineer specializing in machine learning performance optimization. You will optimize models for real-time users, design distributed training strategies, and work with elite ML researchers. Candidates should have at least 8 years of software engineering experience, deep insights into machine learning architectures, and proficiency in PyTorch and NVIDIA optimization. This position offers autonomy in technical decisions and a chance to work with cutting-edge technology.
Qualifications
8+ years of software engineering experience with significant work in ML performance.
Deep insight into modern machine learning architectures.
Track record of owning projects end to end.
Proficiency with PyTorch, Triton, and NVIDIA GPU ecosystems.
Responsibilities
Optimize models for real-time use by hundreds of thousands of users.
Design distributed training strategies for GPU clusters.
Develop tools to identify performance bottlenecks.
Pioneer innovative approaches to enhance performance metrics.
Skills
Software engineering
Machine learning performance optimization
Distributed training
PyTorch
NVIDIA GPU optimization
Tools
Triton
TF/JAX
Job description
Odyssey in Greater London seeks an experienced software engineer specializing in machine learning performance optimization. You will optimize models for real-time users, design distributed training strategies, and work with elite ML researchers. Candidates should have at least 8 years of software engineering experience, deep insights into machine learning architectures, and proficiency in PyTorch and NVIDIA optimization. This position offers autonomy in technical decisions and a chance to work with cutting-edge technology.