Get more replies from employers
Send a job-specific resume in minutes.
Fractile is looking for a Senior ML Runtime Engineer in the UK to integrate AI acceleration hardware with inference frameworks. Join a small expert team to tackle complex problems like KV cache management and scalable multi-user inference.
This hybrid role offers competitive salary and equity, along with a culture that values learning and collaboration. You'll have the opportunity to shape the runtime stack of cutting-edge AI technologies while working in offices located in London and Bristol.
We’re taking a revolutionary approach to computing — building AI acceleration hardware that runs the world’s largest language models 100× faster than existing systems. Our team works at the cutting edge of both hardware and software AI development, and we’re growing fast.
Fractile
We’re looking for a Senior ML Runtime Engineer to help us integrate Fractile’s AI accelerators with the latest inference frameworks and build the runtime stack that makes them fly. You’ll work on genuinely hard problems — KV cache management, scalable multi‑user inference, and the internals of transformer model execution — alongside a collaborative team that values curiosity and rigor equally.
This is a hybrid role, with offices in London and Bristol — your choice of base.
We care most about depth of knowledge and a genuine interest in the problem space. You’ll be a strong fit if you have:
Fractile is committed to building a diverse and inclusive team. We welcome applications from people of all backgrounds and actively encourage candidates from underrepresented groups to apply.