RL Researcher - Post-Training for Omni-Model AI

Nuance Labs

Seattle (WA)

On-site

USD 250,000 - 350,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health savings account contributions
15 days of PTO
Lunch, drinks, and snacks provided

Job summary

Nuance Labs in Seattle is seeking a Member of Technical Staff for RL Research, aimed at recent PhD graduates in AI or ML. In this impactful role, you will own the RL and post-training for large-scale omni models and contribute to developing advanced AI systems.

You will work on building and optimizing the RL/post-training stack, focusing on interactive behavior, emotional response, and real-time conversation quality. Competitive salary plus equity offered.

Qualifications

  • PhD in ML, RL, or a related field, with research publications.
  • Solid understanding of RL/post-training methods.
  • Strong software engineering fundamentals.

Responsibilities

  • Build RL/post-training stack from 0→1 for Nuance.
  • Develop and scale post-training methods.
  • Design systems that connect research ideas to production.

Skills

Machine Learning
Reinforcement Learning
Software Engineering
Research Depth
Model Evaluation

Education

PhD in ML, RL, or related field

Tools

OpenRLHF
vLLM

Job description

Nuance Labs in Seattle is seeking a Member of Technical Staff for RL Research, aimed at recent PhD graduates in AI or ML. In this impactful role, you will own the RL and post-training for large-scale omni models and contribute to developing advanced AI systems.

You will work on building and optimizing the RL/post-training stack, focusing on interactive behavior, emotional response, and real-time conversation quality. Competitive salary plus equity offered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior MTS, RL & Post-Training for Real-Time AI
Senior MTS, RL & Post-Training for Real-Time AI

Nuance Labs • Seattle (WA)

On-site
USD 300,000 - 500,000
Health Savings Account plan
15 days PTO plus holidays
Lunch and snacks provided
Member of Technical Staff — RL Research (Experienced)
Member of Technical Staff — RL Research (Experienced)

Nuance Labs • Seattle (WA)

On-site
USD 300,000 - 500,000
Health Savings Account plan
15 days PTO plus holidays
Lunch and snacks provided
Member of Technical Staff — RL Research (New PhD Grad)
Member of Technical Staff — RL Research (New PhD Grad)

Nuance Labs • Seattle (WA)

On-site
USD 250,000 - 350,000
Health savings account contributions
15 days of PTO
Lunch, drinks, and snacks provided
Research Engineer: RL & Post-Training LLM Systems
Research Engineer: RL & Post-Training LLM Systems

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Member of Technical Staff - Research & Post-training
Member of Technical Staff - Research & Post-training

Preference Model • Seattle (WA)

On-site
USD 200,000 - 350,000
Competitive cash and equity compensation (>90th percentile)
Ownership and autonomy
Health, vision, dental benefits
+4
Research Engineer, RL Environments & Training Infra
Research Engineer, RL Environments & Training Infra

Preference Model • Seattle (WA)

On-site
USD 200,000 - 350,000
Competitive cash and equity compensation (>90th percentile)
Ownership and autonomy
Health, vision, dental benefits
+4
Member of Technical Staff - Research & Post-training
Member of Technical Staff - Research & Post-training

Preference Model • San Francisco (CA)

On-site
USD 120,000 - 150,000
Competitive cash and equity compensation (>90th percentile)
Health, vision, dental benefits
401K match
+2
Post-Training AI Research Engineer – RL & Agentic Infra
Post-Training AI Research Engineer – RL & Agentic Infra

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Agent RL Research Engineer — Pioneering Post-Training AI
Agent RL Research Engineer — Pioneering Post-Training AI

Scale AI, Inc. • New York (NY)

On-site
USD 264,000 - 331,000
Research Engineer, Post-Training
Research Engineer, Post-Training

Storm3 • San Francisco (CA)

On-site
USD 140,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2