Code RL Research Engineer — Build Safe, Scalable AI

Anthropic

San Francisco (CA)

On-site

USD 500,000 - 850,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Anthropic in San Francisco is looking for a Research Engineer to enhance AI systems by advancing their coding abilities—encompassing writing, editing, testing, and debugging real software.

You will design RL tasks and develop reward signals, run training experiments, and play a pivotal role in shaping the future of beneficial AI systems.

The compensation range for this position is USD $500,000 – $850,000 annually.

Qualifications

  • Strong software-engineering skills with deep Python expertise, including async/concurrent programming.
  • Experience owning systems end-to-end and debugging across the stack.
  • Balance research exploration with engineering implementation.

Responsibilities

  • Advance models' ability to write, edit, test, debug, and ship real software.
  • Design RL environments and coding tasks, build the reward signals.
  • Run training experiments on frontier models.

Skills

Software-engineering skills
Deep Python expertise
Debugging across the stack
Knowledge of code quality
Experience with reinforcement learning

Tools

PyTorch
CUDA
TPU

Job description

Anthropic in San Francisco is looking for a Research Engineer to enhance AI systems by advancing their coding abilities—encompassing writing, editing, testing, and debugging real software.

You will design RL tasks and develop reward signals, run training experiments, and play a pivotal role in shaping the future of beneficial AI systems.

The compensation range for this position is USD $500,000 – $850,000 annually.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Code RL Research Engineer — AI Coding & RL Systems
Code RL Research Engineer — AI Coding & RL Systems

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer: AI for Safe, Real-World Software
Research Engineer: AI for Safe, Real-World Software

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
RL Research Engineer - Scalable, Safe AI Systems
RL Research Engineer - Scalable, Safe AI Systems

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+2
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

On-site
USD 500,000 - 850,000
Cybersecurity RL Research Engineer for Safe AI Systems
Cybersecurity RL Research Engineer for Safe AI Systems

Anthropic • San Francisco (CA)

On-site
USD 300,000 - 405,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
+1
Research Engineer
Research Engineer

Confero • San Francisco (CA)

Hybrid
USD 200,000 - 300,000
Research Engineer/Research Scientist, RL/Reasoning
Research Engineer/Research Scientist, RL/Reasoning

Slope • San Francisco (CA)

On-site
USD 310,000 - 460,000
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

OpenAI • San Francisco (CA)

On-site
USD 380,000 - 445,000
Researcher, Synthetic RL
Researcher, Synthetic RL

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 160,000