Stand out for this role — generate a tailored resume and cover letter in about a minute.
Tencent's Lightspeed Tech Center is seeking a PhD-level researcher to advance speech synthesis and multimodal LLMs for real-time spoken interaction. You will prototype and productionize TTS, voice conversion, and audio generation within online applications, collaborating with cross-functional teams to scale experiments to deployment.
Strong background in speech processing, diffusion/autoregressive models, and Python-based DL frameworks is essential; publications at top venues are a plus.
Tencent's Lightspeed Tech Center is seeking a PhD-level researcher to advance speech synthesis and multimodal LLMs for real-time spoken interaction. You will prototype and productionize TTS, voice conversion, and audio generation within online applications, collaborating with cross-functional teams to scale experiments to deployment.
Strong background in speech processing, diffusion/autoregressive models, and Python-based DL frameworks is essential; publications at top venues are a plus.