Staff RL Scientist - Healthcare AI

Hippocratic AI

Menlo Park (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Hippocratic AI is seeking experts to own the Reinforcement Learning (RL) and On-Policy Distillation (OPD) post-training pipeline for healthcare LLMs. You will improve clinical reasoning, safety, and alignment, deploying models to interact with millions of patients across diverse clinical use cases.

The role requires 5+ years in NLP, LLM training, or RL, with 2+ years in RL for LLM post-training, and hands-on experience with large-scale multi-node LLMs in a Palo Alto setting.

Qualifications

  • Requires MS/PhD in CS or related field and strong background in NLP, RL, and ML.
  • Proficiency in Python and PyTorch with experience in large-scale LLM training.
  • Experience with RLHF, RLVR, and post-training methods for LLMs.

Responsibilities

  • Design RL and OPD post-training methods (RLHF, RLVR, OPD) for healthcare LLMs.
  • Build and evaluate reward models, verifiers, and LLM-as-judge pipelines.
  • Develop conversational AI environments and simulations for RL training with synthetic data.
  • Automate post-training loops with agents and automate experiments.
  • Run rigorous experiments to understand drivers of post-training gains.
  • Collaborate with research, engineering, and clinical teams.

Skills

Python
PyTorch
NLP
RL training
RLHF/RLVR

Education

MS or PhD in CS or relevant field

Tools

PyTorch (distributed)

Job description

Hippocratic AI is seeking experts to own the Reinforcement Learning (RL) and On-Policy Distillation (OPD) post-training pipeline for healthcare LLMs. You will improve clinical reasoning, safety, and alignment, deploying models to interact with millions of patients across diverse clinical use cases.

The role requires 5+ years in NLP, LLM training, or RL, with 2+ years in RL for LLM post-training, and hands-on experience with large-scale multi-node LLMs in a Palo Alto setting.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Applied Scientist — Healthcare RL & Safe AI
Staff Applied Scientist — Healthcare RL & Safe AI

Hippocratic-Ai • Menlo Park (CA)

On-site
USD 230,000 - 290,000
Applied Scientist, Reinforcement Learning (Mid, Senior, Staff)
Applied Scientist, Reinforcement Learning (Mid, Senior, Staff)

Hippocratic AI • Menlo Park (CA)

On-site
USD 180,000 - 240,000
Applied Scientist, Reinforcement Learning (Mid, Senior, Staff)
Applied Scientist, Reinforcement Learning (Mid, Senior, Staff)

Hippocratic-Ai • Menlo Park (CA)

On-site
USD 230,000 - 290,000
Remote AI Engineer - Healthcare LLMs & Real-Time AI
Remote AI Engineer - Healthcare LLMs & Real-Time AI

SupportFinity™ • United States

On-site
USD 150,000 - 250,000
$150,000 to $250,000 annual salary.
Meaningful equity stake in the company
Flexible PTO
+2
Remote Senior AI Scientist, NLP/LLM for Healthcare
Remote Senior AI Scientist, NLP/LLM for Healthcare

RXinsider LTD. • United States

Remote
USD 140,000 - 170,000
Competitive salary
Discretionary bonus
Health insurance
+6
Machine Learning Engineer — Self-Improving RL Pipelines
Machine Learning Engineer — Self-Improving RL Pipelines

Hippocratic AI Inc. • Menlo Park (CA)

On-site
USD 180,000 - 240,000
Senior AI/ML Scientist for Healthcare LLMs & GenAI
Senior AI/ML Scientist for Healthcare LLMs & GenAI

Judi Health, LLC • New York (NY)

On-site
USD 180,000 - 226,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Syndesus, Inc. • Austin (TX)

On-site
USD 100,000 - 140,000
100% employer-paid health, vision, and dental insurance
Retirement plans (401(k))
Disability insurance
+1
Senior AI Engineer (LLM Training & RLHF) - Remote
Senior AI Engineer (LLM Training & RLHF) - Remote

Prolific • Virginia Beach (VA)

On-site
USD 100,000 - 140,000
Competitive pay rates
Flexible hours
Ability to work from home
Healthcare AI Scientist: Real-World Data & LLMs
Healthcare AI Scientist: Real-World Data & LLMs

verily • United States

On-site
USD 219,000 - 246,000
Bonus
Equity
Benefits package