Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Spotify is seeking Research Scientists across levels to join its Artist-First AI Music Lab. You’ll conduct groundbreaking research in generative audio, focusing on diffusion or flow matching, vocal synthesis, post-training alignment, and audio editing to create new listening experiences.
You’ll run large-scale experiments with Spotify’s infra, publish findings, and collaborate with engineers, product managers, designers, and researchers to translate ideas into scalable products.
We are seeking Research Scientists (across all levels of seniority) to join our Artist-First AI Music Lab. Our team pioneers and advances state-of-the-art generative technologies for music that create breakthrough experiences for fans and artists. We invent entirely new listening experiences that center and celebrate artists and creatives. All of our products will put artists and songwriters first, through these four principles: Partnerships with record labels, distributors, and music publishers: We’ll develop new products for artists and fans through upfront agreements, not by asking for forgiveness later. Choice in participation: We recognize there’s a wide range of views on the use of generative music tools within the artistic community. Therefore, artists and rightsholders will choose if and how to participate to ensure the use of AI tools aligns with the values of the people behind the music. Fair compensation and new revenue: We will build products that create wholly new revenue streams for rightsholders, artists, and songwriters, ensuring they are properly compensated for the use of their work and transparently credited for their contributions. Artist-fan connection: AI tools we develop will not replace human artistry. They will give artists new ways to be creative and connect with fans. We will leverage our role as the place where more than 700 million people already come to listen to music every month to ensure that generative AI deepens artist-fan connections. Learn more in our press release:
Conduct groundbreaking research in generative audio using diffusion or flow matching models, with a focus on one or more of the following areas:
You will also:
You have a Ph.D. in Computer Science, Mathematics, Engineering, or a related field. Previous industry experience is helpful. You have experience in one or more of the following fields: generative modeling, machine learning, music information retrieval, speech processing, audio processing, signal processing, probabilistic modeling, computer vision, or related areas.
You have deep expertise in at least one of the focus areas above—whether that’s vocal/speech synthesis, post‑training alignment techniques (e.g., PPO, GRPO, DPO), or audio‑to‑audio generation and text‑guided music editing.
You have publications at leading conferences such as ICASSP, ISMIR, INTERSPEECH, ICLR, AAAI, IJCAI, NeurIPS, ICML, CVPR, ECCV, ICCV, or related venues.
You have strong coding skills in Python, PyTorch, and NumPy. You are a creative problem solver who is passionate about building outstanding products that add real value to millions of people. You are enthusiastic about turning research ideas into products operating at scale. You can explain complex topics in simple terms, and you enjoy building strong relationships with colleagues and stakeholders.
We offer you the flexibility to work where you work best! For this role, you can be within the EMEA region as long as we have a work location. This team operates within the Central European and GMT time zone for collaboration. Core working hours are CET 3pm-6pm / EST 9am-12pm.
Spotify is an equal opportunity employer. You are welcome at Spotify for who you are, no matter where you come from, what you look like, or what’s playing in your headphones. Our platform is for everyone, and so is our workplace. The more voices we have represented and amplified in our business, the more we will all thrive, contribute, and be forward-thinking! So bring us your personal…