ML Systems Engineer — RLHF for Safe AI Training

Anthropic Limited

New York (NY)

Hybrid

USD 500,000 - 850,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking an ML Systems Engineer to join the Reinforcement Learning Engineering team. You will work on critical algorithms and infrastructure that researchers rely on to train models like Claude, focusing on speed, reliability, and ease of use.

Ideal candidates have 4+ years of software engineering experience, enjoy working on productive systems, and are interested in ML research and societal impact.

Qualifications

  • Have 4+ years of software engineering experience.
  • Enjoy pair programming and collaborating with others.
  • Experience or interest in machine learning research and RLHF.
  • Interested in building reliable, scalable AI systems with societal impact considerations.

Responsibilities

  • Build, maintain, and improve algorithms and systems used to train AI models.
  • Improve speed, reliability and usability of reinforcement learning training pipelines.
  • Collaborate with researchers to enable faster progress and robust deployment.

Skills

4+ years software engineering
Pair programming
Python
RLHF
Machine learning research
Societal impact awareness

Education

Bachelor’s degree or equivalent

Job description

Anthropic is seeking an ML Systems Engineer to join the Reinforcement Learning Engineering team. You will work on critical algorithms and infrastructure that researchers rely on to train models like Claude, focusing on speed, reliability, and ease of use.

Ideal candidates have 4+ years of software engineering experience, enjoy working on productive systems, and are interested in ML research and societal impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Systems Engineer: RL Training & AI Safety
ML Systems Engineer: RL Training & AI Safety

Anthropic • Seattle (WA)

Hybrid
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
ML Systems Engineer — RL Training & Finetuning
ML Systems Engineer — RL Training & Finetuning

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • Seattle (WA)

Hybrid
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Staff Software Engineer — RL Infrastructure & Platforms
Staff Software Engineer — RL Infrastructure & Platforms

Anthropic • New York (NY), Seattle (WA), San Francisco (CA)

On-site
USD 140,000 - 180,000
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Research Engineer - Large-Scale ML & Safe AI
Research Engineer - Large-Scale ML & Safe AI

Anthropic • New York (NY)

Hybrid
USD 350,000 - 500,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Senior ML Engineer: AI Safety & Alignment (RLHF)
Senior ML Engineer: AI Safety & Alignment (RLHF)

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 400,000
Top-tier salary and equity grants
Comprehensive medical, dental, and eye