Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Hunter Bond is seeking an ML Infrastructure Engineer in London to design and build infrastructure powering high-performance ML workloads. You’ll enable faster training, efficient inference and scalable deployment across distributed environments.
As a key early team member, you’ll own architecture, drive observability and performance benchmarking, and collaborate with ML engineers and platform teams to remove bottlenecks and push the boundaries of model efficiency and throughput.
Salary: up to £200,000 P/A + Bonus (DOE)
Location: London
We are seeking an ML Infrastructure Engineer to join a fast-growing team focused on optimising the performance and scalability of machine learning systems.
In this role, you will design and build the infrastructure that powers high-performance ML workloads, enabling faster training, efficient inference, and scalable deployment across distributed environments. You’ll work on systems that sit at the intersection of machine learning and high-performance engineering, helping to push the boundaries of model efficiency and throughput.
You will collaborate closely with ML engineers, researchers, and platform teams to identify performance bottlenecks and implement optimisations across the stack—from data pipelines and compute orchestration to model execution and hardware utilisation. Your work will directly impact how models are trained, deployed, and scaled in production environments.
As an early member of a growing team, you’ll have significant ownership over architectural decisions, contributing to the design of robust, scalable infrastructure and developer tooling. You’ll also help establish best practices around observability, reliability, and performance benchmarking.
This is an ideal opportunity for someone who enjoys low-level problem solving, distributed systems, and working with modern ML frameworks and infrastructure technologies. You’ll play a key role in building systems that make machine learning faster, more efficient, and more cost-effective.
You will have: