A complete application in a minute — tailored resume and cover letter, ready to send.
RB Labs is building a real-time generative AI system where latency is the core challenge. We’re seeking a model distillation specialist to shrink large models into fast, deployment-ready versions without sacrificing quality.
This is a hands-on role: you’ll design and run distillation experiments, evaluate quality/latency tradeoffs, and ship the results to production GPU infrastructure.
We’re building a real-time generative AI system where latency is the core challenge. We’re looking for a model distillation specialist to help us shrink large models into fast, deployment-ready versions without sacrificing quality.
This is a hands-on role: you’ll design and run distillation experiments, evaluate quality/latency tradeoffs, and ship the results to production GPU infrastructure.
We operate like a research lab: empirical, fast iteration, incomplete docs, self-directed testing. You’ll need to be comfortable designing your own experiments and pushing for more engineering structure as you go.