LLM Alignment & Post-Training Research Scientist

Scale AI

New York (NY)

On-site

USD 150,000 - 200,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health & wellbeing benefits
Learning & development stipend
Community & ERG events
Parental support

Job summary

Scale AI seeks Research Scientists and Research Engineers to advance LLM post-training techniques (SFT, RLHF, reward modeling). The role emphasizes optimizing data curation and algorithms to improve instruction following, factual accuracy, coding, multilingual and multimodal understanding.

You will develop novel methods, collaborate with researchers, and publish in top conferences. Join Scale AI to help shape the next generation of generative models and work with leading labs on strategic inputs

Qualifications

  • Published research in ML at major conferences or journals.
  • Deep understanding of DL, RL, and large-scale model fine-tuning.
  • Excellent written and verbal communication skills.
  • Ph.D. or Master’s in CS/ML/AI or related field.
  • Experience in customer-facing roles.
  • Experience with post-training techniques such as RLHF or instruction tuning.

Responsibilities

  • Develop novel post-training methods to improve LLMs.
  • Collaborate with researchers and engineers to set data-driven best practices.
  • Optimize data curation and algorithms to boost instruction following, factuality, and multilingual/multimodal understanding.
  • Publish findings at top AI conferences and contribute to model robustness.

Skills

Post-training techniques
Reinforcement learning
Deep learning
Communication skills
Customer-facing experience

Education

Ph.D. or Master’s in CS/ML/AI

Job description

Scale AI seeks Research Scientists and Research Engineers to advance LLM post-training techniques (SFT, RLHF, reward modeling). The role emphasizes optimizing data curation and algorithms to improve instruction following, factual accuracy, coding, multilingual and multimodal understanding.

You will develop novel methods, collaborate with researchers, and publish in top conferences. Join Scale AI to help shape the next generation of generative models and work with leading labs on strategic inputs

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Post-Training Research Scientist (SFT & RLHF)
LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Scientist - SFT & RLHF
LLM Post-Training Scientist - SFT & RLHF

Scale AI, Inc. • San Francisco (CA)

On-site
USD 181,000 - 226,000
Health benefits
Retirement plan
Learning stipend
+2
LLM Post-Training Research Scientist
LLM Post-Training Research Scientist

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Research Scientist (SFT/RLHF)
LLM Post-Training Research Scientist (SFT/RLHF)

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
GenAI Research Scientist — LLM Post-Training
GenAI Research Scientist — LLM Post-Training

Scale AI • Seattle (WA)

On-site
USD 166,000 - 207,000
Health coverage
Dental coverage
Vision coverage
+3
Post-Training ML Research Scientist (RLHF/SFT)
Post-Training ML Research Scientist (RLHF/SFT)

United States Digital Space LLC • New York (NY), San Francisco (CA)

On-site
USD 181,000 - 226,000
Health coverage
Equity
Generous PTO
+2
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI • Seattle (WA)

On-site
USD 166,000 - 207,000
Health coverage
Dental coverage
Vision coverage
+3
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2