AI Alignment Research Engineer — RLHF & Human-in-the-Loop

Sterling Inspired Staffing.

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical plans
Generous parental leave
Unlimited PTO
Educational budget
WFH stipend
Daily lunch

Job summary

A leading staffing agency is seeking an Applied Research Engineer to design and build advanced systems for training AI models. This role combines research and engineering, focusing on techniques such as Reinforcement Learning from Human Feedback (RLHF) and improving human data quality. The ideal candidate will have a Ph.D. or Master's in a related field and strong experience in machine learning. Benefits include comprehensive medical plans, unlimited PTO, and educational budget.

Qualifications

  • 3+ years of experience solving complex ML problems with real-world impact.
  • Deep knowledge of frontier model training and alignment techniques.
  • Publication record at top-tier conferences (NeurIPS, ICML, etc.).

Responsibilities

  • Develop state-of-the-art methods for aligning AI systems with human intent.
  • Design systems to measure and improve human feedback quality.
  • Build tools to enhance data labeling processes through AI-assisted workflows.

Skills

Machine Learning problem-solving
Data-centric AI
Human-centered design
Proficiency in Python
Active Learning
Reinforcement Learning techniques

Education

Ph.D. or Masters in Computer Science, Machine Learning, AI, or related field

Tools

PyTorch
JAX
TensorFlow

Job description

A leading staffing agency is seeking an Applied Research Engineer to design and build advanced systems for training AI models. This role combines research and engineering, focusing on techniques such as Reinforcement Learning from Human Feedback (RLHF) and improving human data quality. The ideal candidate will have a Ph.D. or Master's in a related field and strong experience in machine learning. Benefits include comprehensive medical plans, unlimited PTO, and educational budget.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Alignment Research Engineer
AI Alignment Research Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 120,000 - 150,000
Research Engineer - Post-Training & Data Environments
Research Engineer - Post-Training & Data Environments

Mercor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Generous equity grant
$10K housing bonus
$1.5K monthly food stipend
+2
Staff AI Engineer — RL & Post-Training Alignment
Staff AI Engineer — RL & Post-Training Alignment

xAI • Palo Alto (CA)

On-site
USD 180,000 - 600,000
Equity
Comprehensive medical coverage
Vision coverage
+4
AI Alignment Research Scientist Intern — ML Systems
AI Alignment Research Scientist Intern — ML Systems

Meta • Menlo Park (CA)

On-site
USD 80,000 - 100,000
Research Engineer: Multimodal RLHF & Personalized AI
Research Engineer: Multimodal RLHF & Personalized AI

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
Post-Training Research Engineer: ML Alignment & Pipelines
Post-Training Research Engineer: ML Alignment & Pipelines

Character.AI • San Francisco (CA)

On-site
USD 120,000 - 150,000
AI Safety & Alignment Research Scientist
AI Safety & Alignment Research Scientist

Google DeepMind • Mountain View (CA)

On-site
USD 130,000 - 160,000
Agent ML Research Engineer – Post-Training RL, Equity
Agent ML Research Engineer – Post-Training RL, Equity

Scale AI • United States

Hybrid
USD 180,000 - 315,000
Video Generation AI Research Scientist — Alignment & RLHF
Video Generation AI Research Scientist — Alignment & RLHF

Genmo • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior AI Robotics Research Engineer (Humanoid RL)
Senior AI Robotics Research Engineer (Humanoid RL)

Agility Robotics • California (MO)

Hybrid
USD 185,000 - 288,000
401(k) Plan with 6% company match
Company stock options
100% company-paid medical, dental, and vision insurance
+5