LLM Post-Training Scientist - SFT & RLHF

Scale AI, Inc.

San Francisco (CA)

On-site

USD 181,000 - 226,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health benefits
Retirement plan
Learning stipend
Generous PTO
Commuter stipend

Job summary

Scale AI, Inc. seeks Research Scientists and Research Engineers specializing in LLM post-training (SFT, RLHF, reward modeling) to optimize data curation and evaluation for enhancing LLMs in text and multimodal modalities.

You will develop novel methods to improve alignment and generalization of large-scale generative models and collaborate with researchers to define best practices in data-driven AI development.

Qualifications

  • Ph.D. or Master’s in CS/ML/AI required.
  • Deep understanding of deep learning, RL, and large-scale model fine-tuning.
  • Experience with RLHF, preference modeling, or instruction tuning.
  • Excellent written and verbal communication skills.

Responsibilities

  • Research and develop post-training techniques to enhance LLMs.
  • Design and experiment new approaches to preference optimization.
  • Analyze model behavior and propose bias mitigation and robustness solutions.
  • Publish research findings in top-tier AI conferences.

Skills

LLM post-training
Reinforcement learning
Data curation
Eval and bias mitigation
Scientific communication

Education

Ph.D. or Master’s in CS/ML/AI

Job description

Scale AI, Inc. seeks Research Scientists and Research Engineers specializing in LLM post-training (SFT, RLHF, reward modeling) to optimize data curation and evaluation for enhancing LLMs in text and multimodal modalities.

You will develop novel methods to improve alignment and generalization of large-scale generative models and collaborate with researchers to define best practices in data-driven AI development.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Post-Training Research Scientist (SFT & RLHF)
LLM Post-Training Research Scientist (SFT & RLHF)

Scale AI, Inc. • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental & vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Research Scientist (SFT/RLHF)
LLM Post-Training Research Scientist (SFT/RLHF)

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
Generative AI Research Scientist: LLM Post-Training
Generative AI Research Scientist: LLM Post-Training

Scale AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
LLM Post-Training Research Scientist
LLM Post-Training Research Scientist

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+2
Post-Training ML Research Scientist (RLHF/SFT)
Post-Training ML Research Scientist (RLHF/SFT)

United States Digital Space LLC • New York (NY), San Francisco (CA)

On-site
USD 181,000 - 226,000
Health coverage
Equity
Generous PTO
+2
LLM Evaluation Scientist — Benchmarks & Failure Analysis
LLM Evaluation Scientist — Benchmarks & Failure Analysis

Scale AI, Inc. • Seattle (WA)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning and development stipend
+3
Applied Post-Training LLM Research Scientist
Applied Post-Training LLM Research Scientist

Modal Labs • New York (NY)

On-site
USD 180,000 - 240,000
LLM Post-Training Engineer - RL, SFT & Data Pipelines
LLM Post-Training Engineer - RL, SFT & Data Pipelines

GoTo Meeting • Mountain View (CA)

On-site
USD 150,000 - 230,000
Health, dental, and vision care for you and your family
Top-tier 401(K) plan with company matching
Paid time off and paid holidays
+2
Machine Learning Research Scientist, Post-Training
Machine Learning Research Scientist, Post-Training

Scale AI • New York (NY)

On-site
USD 181,000 - 226,000
Health, dental and vision coverage
Retirement benefits
Learning & development stipend
+2
RL Research Scientist for LLM Post-Training & Code Models
RL Research Scientist for LLM Post-Training & Code Models

AMD • Santa Clara (CA)

On-site
USD 180,000 - 260,000
AMD benefits