An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Desert Ant Labs in Amsterdam offers a hands-on role to advance speech, vision, video, and text models for on-device deployment. You will help shrink models to fit real devices and ship them in our Detail and Subwave apps.
The role covers data generation, benchmarking across languages, and collaborating with teams to define model capabilities and evaluation standards. Remote work options exist alongside Amsterdam-based collaboration.
Train our speech, vision, video, and text models, make them smaller and faster on real devices, and ship them in Detail and Subwave, our own apps.
Location Amsterdam, Remote
Time zone UTC-5 to UTC+1
Type Full time
We pick each model's default settings for the products that use the model, such as which kinds of personal data Redact removes when the developer changes nothing, or whether Uhm leaves a filler word in or risks cutting a real word. You make those decisions with the team, and we expect you to say so when you disagree.
We want a published benchmark for every language a model supports, and the same quality in each. We aren't there yet for every model, and closing those gaps is part of the work.
Start what needs starting without waiting to be asked, and finish what you start. Take on work outside your role when a project needs you.
We ship quickly, so we cut scope until only the part users notice is left. Anyone can comment on your work or redo your draft, and we say early when work isn't ready. We read that feedback as help.
We judge the work by what shipped and what changed because of it. Nobody counts hours.