LLM Post-Training Research Scientist (SFT/RLHF)

Scale AI

New York (NY)

On-site

USD 181,000 - 226,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health, dental and vision coverage
Retirement benefits
Learning & development stipend
Generous PTO
Commuter stipend

Job summary

Scale AI seeks Research Scientists and Research Engineers in New York to advance post-training methods for LLMs, including SFT and RLHF, across text and multimodal data. You will collaborate with researchers to refine data curation and evaluation practices and contribute to the next generation of generative models.

The role emphasizes rigorous analysis of model behavior, development of bias mitigation strategies, and publication of results in leading AI venues.

Qualifications

  • PhD or MSc in CS/ML or related field
  • Strong background in deep learning, RL, and large-scale model fine-tuning
  • Experience with post-training techniques such as RLHF
  • Excellent written and verbal communication
  • Published research at major ML conferences or journals
  • Previous customer-facing experience

Responsibilities

  • Research and develop post-training techniques to enhance LLMs in text and multimodal modalities
  • Design and experiment new approaches to preference optimization
  • Analyze model behavior, identify weaknesses, and propose bias mitigation and robustness improvements
  • Publish research findings in top-tier AI conferences

Skills

SFT
RLHF
Reward modeling
Bias mitigation
Model evaluation
Publications
Communication
Customer facing

Education

PhD or MSc in CS/ML

Job description

Scale AI seeks Research Scientists and Research Engineers in New York to advance post-training methods for LLMs, including SFT and RLHF, across text and multimodal data. You will collaborate with researchers to refine data curation and evaluation practices and contribute to the next generation of generative models.

The role emphasizes rigorous analysis of model behavior, development of bias mitigation strategies, and publication of results in leading AI venues.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Post-Training Research Scientist (SFT & RLHF)
LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Scientist - SFT & RLHF
LLM Post-Training Scientist - SFT & RLHF

Scale AI, Inc. • San Francisco (CA)

On-site
USD 181,000 - 226,000
Health benefits
Retirement plan
Learning stipend
+2
Post-Training ML Research Scientist (RLHF/SFT)
Post-Training ML Research Scientist (RLHF/SFT)

United States Digital Space LLC • New York (NY), San Francisco (CA)

On-site
USD 181,000 - 226,000
Health coverage
Equity
Generous PTO
+2
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Research Scientist
LLM Post-Training Research Scientist

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Production AI Researcher - LLMs & Agents for SRE
Production AI Researcher - LLMs & Agents for SRE

AI • New York (NY)

On-site
USD 160,000 - 300,000
Health insurance
Startup equity
Flexible time off
+1
LLM Fine-Tuning Engineer for Production ML
LLM Fine-Tuning Engineer for Production ML

scalr • New York (NY)

On-site
USD 200,000 - 300,000
LLM Evaluation Scientist — Benchmarks & Failure Analysis
LLM Evaluation Scientist — Benchmarks & Failure Analysis

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+3
LLM Evaluation Scientist: Benchmarks & Failure Insights
LLM Evaluation Scientist: Benchmarks & Failure Insights

AI Chopping Block • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health coverage
Dental coverage
Vision coverage
+4