Applied Researcher, Generative Audio Systems

Cartesia

San Francisco (CA)

On-site

USD 180,000 - 260,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive salary
Equity package
Health insurance
Parental leave
401(k)
Commuter allowance
Flexible PTO
Meals & snacks
Your own personal Yoshi

Job summary

Cartesia is hiring for an Audio Post-Training role to build and improve capabilities that shape how users interact with our generative audio models. This team spans research to production, designing evaluations, data pipelines, finetuning, RL, and model evaluation.

You will translate customer needs into research plans, own end-to-end delivery, and communicate improvements to product and customers. A strong foundation in software and ML, plus a drive to solve real-world problems, is essential.

Qualifications

  • Strong fundamentals in software engineering, machine learning, debugging complex systems, and the ability + desire to learn quickly.
  • Experience building and ensuring quality of large multilingual datasets.
  • Experience training and debugging generative models (speech, text, or multimodal), especially SFT, RL, synthetic data, and evaluation (both human and automated).
  • Excitement about solving problems grounded in real customer needs, not just benchmarks.
  • Bonus points if you have native proficiency in other languages!

Responsibilities

  • Collaborate with product teams to understand and prioritize customer asks
  • Cut through the ambiguity of vaguely described behavioral problems to make concrete research plans.
  • Ideate and experiment across the full modeling stack, including data processing, synthetic data, SFT, RL, and evals to solve for high priority model capabilities
  • Root cause failures in production models and understand how to fix them in future model iterations
  • Decide which features and capabilities are ready for public launch

Skills

Software engineering
Machine learning
Debugging complex systems
Multilingual datasets
Generative models
SFT / RL experience
Customer-focused problem solving
Multilingual proficiency bonus

Job description

Cartesia is hiring for an Audio Post-Training role to build and improve capabilities that shape how users interact with our generative audio models. This team spans research to production, designing evaluations, data pipelines, finetuning, RL, and model evaluation.

You will translate customer needs into research plans, own end-to-end delivery, and communicate improvements to product and customers. A strong foundation in software and ML, plus a drive to solve real-world problems, is essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Applied Researcher, Audio Post-Training
Applied Researcher, Audio Post-Training

Cartesia • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive salary
Equity package
Health insurance
+6
Lead Audio AI Research for Multimodal Systems
Lead Audio AI Research for Multimodal Systems

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Remote Research Scientist - Generative Audio & ML
Remote Research Scientist - Generative Audio & ML

Spotify • New York (NY)

On-site
USD 133,000 - 191,000
Health insurance
Six month paid parental leave
401(k) retirement plan
+3
Staff Research Scientist (Diffusion)
Staff Research Scientist (Diffusion)

Fabrik Talent • United States

Remote
GBP 80,000 - 120,000
Remote Research Engineer — AI Audio & Models
Remote Research Engineer — AI Audio & Models

ElevenLabs • Town of Poland (NY)

On-site
USD 140,000 - 190,000
Staff Research Scientist, Audio Diffusion (Remote)
Staff Research Scientist, Audio Diffusion (Remote)

Fabrik Talent • United States

Remote
GBP 80,000 - 120,000
Lead Data Engineer, Generative Audio Training Data
Lead Data Engineer, Generative Audio Training Data

Adobe Inc. • San Francisco (CA)

On-site
USD 216,000 - 313,000
Remote Audio AI Research Engineer: LLM & Reasoning
Remote Audio AI Research Engineer: LLM & Reasoning

Centific Global Solutions, Inc. • Redmond (WA)

Hybrid
USD 150,000 - 160,000
Competitive compensation
Hybrid/Remote options
GPU infrastructure access
+1
Research, Audio Expertise
Research, Audio Expertise

Mosaic.tech • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3
Applied Audio ML Engineer: Post-Training & On-Device
Applied Audio ML Engineer: Post-Training & On-Device

Liquid AI • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive base salary
Equity in a unicorn-stage company
100% paid health premiums
+2