ML Systems Engineer — RL Training & Finetuning

Anthropic

San Francisco (CA)

Hybrid

USD 500,000 - 850,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive compensation
Equity donation matching
Generous vacation and parental leave
Flexible working hours

Job summary

Anthropic is seeking an ML Systems Engineer to join the Reinforcement Learning Engineering team in San Francisco. You will develop critical algorithms and infrastructure to train AI models, focusing on performance, robustness, and usability to accelerate research progress.

You’ll work with finetuning researchers using RLHF and related methods, building systems that empower researchers to train and evaluate models safely and efficiently in a fast-paced environment.

Qualifications

  • 4+ years of software engineering experience.
  • ey Bachelor’s degree or higher in a related field.
  • Interest in ML research and safety and societal impact.

Responsibilities

  • Build, maintain and improve RL algorithms and infra used by researchers.
  • Improve speed, reliability, and usability of reinforcement learning systems.
  • Support and empower research teams training production Claude models.

Skills

4+ years software engineering
Distributed systems
Python
RLHF
Pair programming

Education

Bachelor’s degree

Job description

Anthropic is seeking an ML Systems Engineer to join the Reinforcement Learning Engineering team in San Francisco. You will develop critical algorithms and infrastructure to train AI models, focusing on performance, robustness, and usability to accelerate research progress.

You’ll work with finetuning researchers using RLHF and related methods, building systems that empower researchers to train and evaluate models safely and efficiently in a fast-paced environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Systems Engineer — RLHF for Safe AI Training
ML Systems Engineer — RLHF for Safe AI Training

Anthropic Limited • New York (NY)

Hybrid
USD 500,000 - 850,000
ML Systems Engineer: RL Training & AI Safety
ML Systems Engineer: RL Training & AI Safety

Anthropic • Seattle (WA)

Hybrid
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Staff AI Systems Engineer — Inference & RL
Staff AI Systems Engineer — Inference & RL

Together • San Francisco (CA)

On-site
USD 200,000 - 280,000
Health insurance
Startup equity
Competitive benefits
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
Senior ML Systems Engineer: Training Infra
Senior ML Systems Engineer: Training Infra

Neura Market • San Francisco (CA)

On-site
USD 295,000 - 380,000
Relocation assistance
Staff Software Engineer — RL Infrastructure & Platforms
Staff Software Engineer — RL Infrastructure & Platforms

Anthropic • New York (NY), Seattle (WA), San Francisco (CA)

On-site
USD 140,000 - 180,000
RL Infrastructure Engineer — Scalable Training & Performance
RL Infrastructure Engineer — Scalable Training & Performance

xAI • Palo Alto (CA)

On-site
USD 170,000 - 260,000
Health insurance
Life and AD&D insurance
Fertility benefits
+3
Staff Engineer - RL Training Infrastructure
Staff Engineer - RL Training Infrastructure

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
AI Systems Engineer — RL Environments & Scalable Infra
AI Systems Engineer — RL Environments & Scalable Infra

AI Talent Now • San Francisco (CA)

On-site
USD 120,000 - 150,000