Senior MTS, RL & Post-Training for Real-Time AI

Nuance Labs

Seattle (WA)

On-site

USD 300,000 - 500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health Savings Account plan
15 days PTO plus holidays
Lunch and snacks provided

Job summary

Nuance Labs in Seattle is looking for a deeply technical Member of Technical Staff to own reinforcement learning and post-training for large models. This role requires significant experience in RL and post-training methods, ensuring the development of systems that achieve real-time, multimodal interaction.

The candidate will build and optimize training systems, implementing modern methods and aligning research with production needs. A focus on interaction quality and feedback loops for omni behavior is crucial.

Compensation ranges from $300,000 to $500,000 base salary, with equity and other benefits.

Qualifications

  • Significant hands-on experience with RL, RLHF, or large-scale fine-tuning for models.
  • Deep understanding of RL/post-training methods and evaluation.
  • Track record of reasoning about model behavior and training dynamics.

Responsibilities

  • Build Nuance's RL/post-training stack from 0→1.
  • Develop and scale post-training methods such as PPO and RLHF.
  • Optimize the end-to-end post-training loop for real-time interaction.

Skills

Reinforcement Learning (RL)
Post-training methods
Policy optimization
Reward modeling
Software engineering fundamentals

Tools

OpenRLHF
vLLM

Job description

Nuance Labs in Seattle is looking for a deeply technical Member of Technical Staff to own reinforcement learning and post-training for large models. This role requires significant experience in RL and post-training methods, ensuring the development of systems that achieve real-time, multimodal interaction.

The candidate will build and optimize training systems, implementing modern methods and aligning research with production needs. A focus on interaction quality and feedback loops for omni behavior is crucial.

Compensation ranges from $300,000 to $500,000 base salary, with equity and other benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Researcher - Post-Training for Omni-Model AI
RL Researcher - Post-Training for Omni-Model AI

Nuance Labs • Seattle (WA)

On-site
USD 250,000 - 350,000
Health savings account contributions
15 days of PTO
Lunch, drinks, and snacks provided
Member of Technical Staff — RL Research (Experienced)
Member of Technical Staff — RL Research (Experienced)

Nuance Labs • Seattle (WA)

On-site
USD 300,000 - 500,000
Health Savings Account plan
15 days PTO plus holidays
Lunch and snacks provided
Real-time Multimodal RL Engineer - Post-Training
Real-time Multimodal RL Engineer - Post-Training

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
Member of Technical Staff, Post-training
Member of Technical Staff, Post-training

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Senior Staff, Multimodal RL & Real-Time AI
Senior Staff, Multimodal RL & Real-Time AI

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
New Grad Staff Engineer — Real-Time Model Optimization
New Grad Staff Engineer — Real-Time Model Optimization

Nuance Labs • Seattle (WA)

On-site
USD 200,000 - 300,000
Health Savings Account with $2,000 annual contributions
15 days of PTO plus public holidays
Lunch, drinks, and snacks provided daily
Member of Technical Staff — RL Research (New PhD Grad)
Member of Technical Staff — RL Research (New PhD Grad)

Nuance Labs • Seattle (WA)

On-site
USD 250,000 - 350,000
Health savings account contributions
15 days of PTO
Lunch, drinks, and snacks provided
Member of Technical Staff, Post-training San Jose
Member of Technical Staff, Post-training San Jose

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
Member of Technical Staff, Mid-training San Jose
Member of Technical Staff, Mid-training San Jose

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
ML Systems Engineer – End-to-End Inference & RL
ML Systems Engineer – End-to-End Inference & RL

Togetherai • San Francisco (CA)

On-site
USD 200,000 - 280,000
Startup equity
Health insurance
Competitive benefits