LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc.

New York (NY)

On-site

USD 181,000 - 226,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health, dental & vision coverage
Retirement benefits
Learning and development stipend
Generous PTO
Commuter stipend

Job summary

Scale AI, Inc. is seeking Research Scientists and Research Engineers to advance LLM post-training techniques, including SFT, RLHF, and reward modeling, with a focus on data curation and evaluation to improve performance in text and multimodal modalities.

You will collaborate with researchers and engineers to define best practices in data-driven AI development and contribute to the next generation of generative models.

Qualifications

  • Ph.D. or Master’s in Computer Science, Machine Learning, AI, or a related field.
  • Deep understanding of deep learning, reinforcement learning, and large-scale model fine-tuning.
  • Experience with post-training techniques such as RLHF, preference modeling, or instruction tuning.
  • Excellent written and verbal communication skills.

Responsibilities

  • Research and develop novel post-training techniques to enhance LLM core capabilities in text and multimodal modalities.
  • Design and experiment new approaches to preference optimization.
  • Analyze model behavior, identify weaknesses, and propose solutions for bias mitigation and model robustness.
  • Publish research findings in top-tier AI conferences.

Skills

SFT
RLHF
Reward modeling
Fine-tuning large models
Reinforcement learning
Multimodal data

Education

Ph.D./Master’s in CS/ML/AI

Job description

Scale AI, Inc. is seeking Research Scientists and Research Engineers to advance LLM post-training techniques, including SFT, RLHF, and reward modeling, with a focus on data curation and evaluation to improve performance in text and multimodal modalities.

You will collaborate with researchers and engineers to define best practices in data-driven AI development and contribute to the next generation of generative models.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Post-Training Scientist - SFT & RLHF
LLM Post-Training Scientist - SFT & RLHF

Scale AI, Inc. • San Francisco (CA)

On-site
USD 181,000 - 226,000
Health benefits
Retirement plan
Learning stipend
+2
LLM Post-Training Research Scientist (SFT/RLHF)
LLM Post-Training Research Scientist (SFT/RLHF)

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
Post-Training ML Research Scientist (RLHF/SFT)
Post-Training ML Research Scientist (RLHF/SFT)

United States Digital Space LLC • New York (NY), San Francisco (CA)

On-site
USD 181,000 - 226,000
Health coverage
Equity
Generous PTO
+2
LLM Post-Training Research Scientist
LLM Post-Training Research Scientist

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Evaluation Scientist — Benchmarks & Failure Analysis
LLM Evaluation Scientist — Benchmarks & Failure Analysis

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+3
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI, Inc. • San Francisco (CA)

On-site
USD 181,000 - 226,000
Health benefits
Retirement plan
Learning stipend
+2