Senior Speech AI Scientist — Generative Voice

Spotify

New York (NY)

Hybrid

USD 169,000 - 242,000

Full time

8 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Spotify is seeking a senior applied research scientist in New York City to lead developments in state-of-the-art speech models and contribute to end-to-end production pipelines.

You will collaborate with engineering and data teams to advance speech synthesis/recognition, explore new ideas, and push the frontiers of speech technology while working in a hybrid environment with some in-person meetings.

Qualifications

  • PhD in ML with professional research experience.
  • Experience with transformers, GANs, diffusion models, flow matching, VAEs, or audio codecs.
  • Experience developing generative models for speech synthesis/recognition, audio/music, NLP, or computer vision.
  • Strong Python and PyTorch skills.

Responsibilities

  • Develop and experiment with new methods for speech synthesis and recognition.
  • Expand speech use-cases across markets and products.
  • Collaborate with engineering/data teams to build scalable pipelines.
  • Share research findings with other researchers in Speak.
  • Help turn ideas into scalable products.

Skills

ML expertise
Python
PyTorch
Research experience
Communication skills

Education

PhD in ML or related field

Tools

Transformers
GANs
Diffusion models
Flow matching
VAEs
Audio codecs

Job description

Spotify is seeking a senior applied research scientist in New York City to lead developments in state-of-the-art speech models and contribute to end-to-end production pipelines.

You will collaborate with engineering and data teams to advance speech synthesis/recognition, explore new ideas, and push the frontiers of speech technology while working in a hybrid environment with some in-person meetings.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Applied Research Scientist - Generative Speech AI
Senior Applied Research Scientist - Generative Speech AI

Spotify • New York (NY)

Hybrid
USD 169,000 - 242,000
Health insurance
Six month paid parental leave
401(k) retirement plan
+3
Generative Audio Research Scientist — Vocal Synthesis (Equity)
Generative Audio Research Scientist — Vocal Synthesis (Equity)

Spotify AB • New York (NY)

On-site
USD 133,000 - 190,000
Health insurance
Six months parental leave
401(k) retirement plan
+3
Pioneering Generative Audio Scientist
Pioneering Generative Audio Scientist

Creandum Advisor LLP • New York (NY)

On-site
USD 133,194 - 190,278
Health insurance
Equity
Paid parental leave
Remote Research Scientist - Generative Audio & ML
Remote Research Scientist - Generative Audio & ML

Spotify • New York (NY)

On-site
USD 133,194 - 190,278
Health insurance
Six month paid parental leave
401(k) retirement plan
+3
Senior Audio & Speech AI Scientist
Senior Audio & Speech AI Scientist

Adobe Inc. • San Francisco (CA)

On-site
USD 187,000 - 271,000
Senior Applied Research Scientist - Personalization
Senior Applied Research Scientist - Personalization

Spotify • New York (NY)

On-site
USD 169,000 - 242,000
Health insurance
Six month paid parental leave
401(k) retirement plan
+3
Research Scientist - Speech
Research Scientist - Speech

JAM • United States

On-site
USD 140,000 - 230,000
Research Scientist - Speech
Research Scientist - Speech

JAM • United States

On-site
USD 100,000 - 130,000
Senior AI Music Research Scientist
Senior AI Music Research Scientist

SupportFinity™ • New York (NY)

Hybrid
USD 164,000 - 235,000
Health insurance
Six month paid parental leave
401(k) retirement plan
+4
Senior AI Voice Quality Analyst
Senior AI Voice Quality Analyst

Spotify AB • New York (NY)

Hybrid
USD 114,000 - 164,000
Health insurance
Six-month paid parental leave
401(k) retirement plan
+4