Get more replies from employers
Send a job-specific resume in minutes.
DeepRec.ai is a fast-growing voice AI company in SF, delivering production-ready AI phone agents. This full-time role focuses on building end-to-end speech technologies across TTS, STT, and neural codecs, with a goal of real-time, human-like interactions for enterprise customers.
You will push theory to production, train on massive audio datasets, and collaborate with engineering and product teams to ship capabilities to customers quickly. A PhD is welcome but not required.
Full-time / Permanent
DeepRec has partnered with a fast-growing, revenue-generating voice AI company empowering enterprises to build AI phone agents at scale. Recent Series C funding with backing from leading Silicon Valley investors, they are building the models and infrastructure that make voice the primary interface between businesses and their customers.
This company has built all of their current models completely in-house, and every model ships to real, paying customers almost immediately. No speculative research track here. If you want your work to hit production within weeks, not years, this is that role.
The Opportunity
The research team are working toward a single, ambitious goal: a fully speech-to-speech conversational AI model that understands and responds like a human, in real time. You'll work across the core building blocks of that roadmap, such as: speech-to-text, text-to-speech, neural audio codecs, and getting LLMs to understand and reason over audio directly.
You'll take ideas from theory through large-scale training to production inference serving millions of calls a day, working closely with engineering and product teams to get your research into real customer environments fast.
What You'll Do
What You'll Bring
We encourage you to apply even if you don't meet every requirement. The right mindset and genuine curiosity matter as much as the resume.
What's In It For You