Senior ML Engineer - AI Alignment & RLHF (Remote)

AI Breaking Wire

San Francisco, Northern (CA, KY)

Hybrid

USD 180,000 - 260,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Equity compensation
Competitive salary
Remote and hybrid options
Catered lunches and wellness stipends

Job summary

Anthropic is seeking a Senior Machine Learning Engineer to advance Constitutional AI and model alignment. You will scale alignment pipelines to keep Claude safe, helpful, and honest as capabilities grow.

Join a team partnering with researchers to translate alignment theory into production training loops, develop scalable RLHF systems, and build automated evaluations for biases and vulnerabilities.

The role offers flexible remote/hybrid work and competitive compensation.

Qualifications

  • 4+ years of industry experience building, training, and deploying large-scale ML models.
  • Strong experience with PyTorch and distributed training infrastructure (Megatron-LM, DeepSpeed, or similar).
  • Solid foundation in alignment methods, RLHF, and automated red-teaming.
  • Passion for AI safety and responsible development of AGI.

Responsibilities

  • Build and optimize scalable RLHF and constitutional AI training loops.
  • Develop automated evaluation frameworks to detect model vulnerabilities, bias, and undesirable behaviors.
  • Partner with research scientists to translate alignment techniques into production-grade training systems.

Skills

RLHF experience
PyTorch
Distributed training
Alignment methodologies

Tools

Megatron-LM
DeepSpeed

Job description

Anthropic is seeking a Senior Machine Learning Engineer to advance Constitutional AI and model alignment. You will scale alignment pipelines to keep Claude safe, helpful, and honest as capabilities grow.

Join a team partnering with researchers to translate alignment theory into production training loops, develop scalable RLHF systems, and build automated evaluations for biases and vulnerabilities.

The role offers flexible remote/hybrid work and competitive compensation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
Senior Machine Learning Engineer, Alignment and Safety
Senior Machine Learning Engineer, Alignment and Safety

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Equity compensation
Competitive salary
Remote and hybrid options
+1
Senior Machine Learning Engineer, Safety & Alignment
Senior Machine Learning Engineer, Safety & Alignment

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
Research Scientist, AI Alignment & Safety
Research Scientist, AI Alignment & Safety

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
ML Systems Engineer: RL Training & AI Safety
ML Systems Engineer: RL Training & AI Safety

Anthropic • Seattle (WA)

Hybrid
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Scientist, Alignment
Research Scientist, Alignment

AI Breaking Wire • San Francisco (CA)

On-site
USD 300,000 - 450,000
Competitive salary and equity packages
Comprehensive health, dental, and visa
Flexible working arrangements and PTO
+1
Research Engineer — AI Safety & Alignment (Bay Area)
Research Engineer — AI Safety & Alignment (Bay Area)

Anthropic • California (MO)

Hybrid
USD 350,000 - 500,000
Senior AI Engineer (LLM Training & RLHF) - Remote
Senior AI Engineer (LLM Training & RLHF) - Remote

Prolific • Virginia Beach (VA)

On-site
USD 100,000 - 140,000
Competitive pay rates
Flexible hours
Ability to work from home
RL Systems Engineer: Build Safe, Steerable AI
RL Systems Engineer: Build Safe, Steerable AI

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer / Scientist, Alignment
Research Engineer / Scientist, Alignment

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 500,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours