Research Scientist - Real-Time Multimodal AI Avatars

Nuance Labs

Seattle (WA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

An innovative AI startup in Seattle is looking for experts to develop a groundbreaking human foundation model that integrates text, speech, and emotional signals in real time. Candidates should have a PhD or equivalent experience with a strong background in training audio generation models and deep learning. You will work alongside a top-tier research team to create lifelike avatars that understand and respond to human nuances, helping bridge the emotional gap in AI. This role offers a collaborative and fast-paced environment where your contributions will directly impact the development of cutting-edge technology.

Qualifications

  • PhD (or equivalent) with experience in training audio generation models.
  • Strong understanding of deep learning and ML pipeline.
  • Ability to solve blank-page problems independently.

Responsibilities

  • Develop the first human foundation model integrating text, speech, and body language.
  • Create lifelike avatars capable of nuanced responses.
  • Innovate in real-time multimodal interaction technology.

Skills

Training speech synthesis models
Deep learning
ML pipeline management
Clean code writing
Collaboration with diverse teams

Education

PhD or equivalent experience in relevant fields

Job description

An innovative AI startup in Seattle is looking for experts to develop a groundbreaking human foundation model that integrates text, speech, and emotional signals in real time. Candidates should have a PhD or equivalent experience with a strong background in training audio generation models and deep learning. You will work alongside a top-tier research team to create lifelike avatars that understand and respond to human nuances, helping bridge the emotional gap in AI. This role offers a collaborative and fast-paced environment where your contributions will directly impact the development of cutting-edge technology.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Conversational Modelling Research Engineer
Conversational Modelling Research Engineer

Tavus • United States

Hybrid
USD 100,000 - 150,000
Multimodal Conversational Research Engineer
Multimodal Conversational Research Engineer

Tavus • United States

Hybrid
USD 100,000 - 150,000
Conversational Modelling Research Engineer
Conversational Modelling Research Engineer

Tavus • San Francisco (CA)

Hybrid
USD 120,000 - 230,000
Research Scientist - Video Diffusion
Research Scientist - Video Diffusion

Nuance Labs • Seattle (WA)

On-site
USD 150,000 - 210,000
Research Assistant
Research Assistant

Nuance Labs • Seattle (WA)

Hybrid
USD 28,000 - 41,000
Remote Senior AI Researcher: Multimodal Foundation Models
Remote Senior AI Researcher: Multimodal Foundation Models

hum.ai • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Senior Multimodal AI Researcher — Flexible, Immersive Media
Senior Multimodal AI Researcher — Flexible, Immersive Media

Dolby Laboratories • Atlanta (GA)

On-site
USD 140,000 - 170,000
Lead Real-Time Multimodal AI Scientist
Lead Real-Time Multimodal AI Scientist

Amazon • Sunnyvale (CA)

On-site
USD 229,000 - 309,000
Lead Real-Time Multimodal AI Scientist, Conversational/AGI
Lead Real-Time Multimodal AI Scientist, Conversational/AGI

Amazon • Bellevue (WA)

On-site
USD 180,000 - 260,000
Research Scientist, Speech & Audio Gen & Multimodal AI
Research Scientist, Speech & Audio Gen & Multimodal AI

Lightspeed Studios • Bellevue (WA)

On-site
USD 122,000 - 230,000
Sign‑on payment
Relocation package
Medical, dental, and vision benefits
+2