ML Systems Engineer: RL Training & AI Safety

Anthropic

Seattle (WA)

Hybrid

USD 520,000 - 850,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity donation matching
Generous vacation
Parental leave
Flexible working hours
Office space

Job summary

Anthropic is seeking an ML Systems Engineer on the Reinforcement Learning Engineering team to build the algorithms and infrastructure used to train Claude-like models. You will focus on performance, robustness, and usability to accelerate research progress and model capabilities.

You will work on production RLHF training systems, improve tooling, and collaborate with researchers. This role blends research support with engineering ownership, in a hybrid office setting.

Qualifications

  • 4+ years of software engineering experience.
  • Like working on systems and tools that make other people more productive.
  • Are results-oriented, with a bias towards flexibility and impact.
  • Pair programming is valued and practiced.
  • Interest in machine learning research.
  • Care about the societal impacts of your work.

Responsibilities

  • Build, maintain, and improve algorithms and systems that researchers use to train models.
  • Improve speed, reliability, and ease-of-use of these systems.
  • Support and empower our research team to progress quickly and safely.
  • Profiling and optimizing RL training pipelines and infrastructure.

Skills

4+ years software engineering
Pair programming
Results oriented
Interest in ML research
Societal impact
Productivity tooling

Education

Bachelor’s degree or equivalent

Tools

Python

Job description

Anthropic is seeking an ML Systems Engineer on the Reinforcement Learning Engineering team to build the algorithms and infrastructure used to train Claude-like models. You will focus on performance, robustness, and usability to accelerate research progress and model capabilities.

You will work on production RLHF training systems, improve tooling, and collaborate with researchers. This role blends research support with engineering ownership, in a hybrid office setting.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Systems Engineer — RLHF for Safe AI Training
ML Systems Engineer — RLHF for Safe AI Training

Anthropic Limited • New York (NY)

Hybrid
USD 500,000 - 850,000
ML Systems Engineer — RL Training & Finetuning
ML Systems Engineer — RL Training & Finetuning

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Senior Software Engineer - End-to-End AI Data & RL
Senior Software Engineer - End-to-End AI Data & RL

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 320,000 - 485,000
Staff Engineer - RL Training Infrastructure
Staff Engineer - RL Training Infrastructure

Pantera Capital • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye
Staff Software Engineer — RL Infrastructure & Platforms
Staff Software Engineer — RL Infrastructure & Platforms

Anthropic • New York (NY), Seattle (WA), San Francisco (CA)

On-site
USD 140,000 - 180,000
ML Systems Engineer for RL & Inference Infrastructure
ML Systems Engineer for RL & Inference Infrastructure

Advanced Micro Devices • Santa Clara (CA)

Hybrid
USD 160,000 - 210,000
AMD benefits
RL Systems Engineer: Inference & Training at Scale
RL Systems Engineer: Inference & Training at Scale

xAI • Palo Alto (CA)

On-site
USD 180,000 - 240,000