Code RL Research Engineer — AI Coding & RL Systems

Anthropic

San Francisco (CA)

Hybrid

USD 500,000 - 850,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking a Research Engineer for the Code RL team in San Francisco. This role involves designing RL environments, running training experiments, and improving model training pipelines. Candidates should have strong software engineering skills, deep Python expertise, and a commitment to building safe AI systems.

The ideal candidate will balance research exploration with engineering implementation, ensuring high code quality and performance. Salary range is between $500,000 to $850,000 annually.

Qualifications

  • Strong software-engineering skills and deep Python expertise, including async and concurrent programming.
  • Experience with reinforcement learning, RLHF, post-training, or LLM finetuning.
  • Experience with large-scale distributed training and performance profiling.

Responsibilities

  • Design RL environments and coding tasks, and build reward signals.
  • Run training experiments on frontier models and diagnose model behavior.
  • Improve speed and reliability of training pipelines.

Skills

Software engineering skills
Deep Python expertise
Debugging across the stack
Research exploration and engineering implementation

Education

Bachelor’s degree or equivalent

Tools

PyTorch
CUDA/GPU or TPU

Job description

Anthropic is seeking a Research Engineer for the Code RL team in San Francisco. This role involves designing RL environments, running training experiments, and improving model training pipelines. Candidates should have strong software engineering skills, deep Python expertise, and a commitment to building safe AI systems.

The ideal candidate will balance research exploration with engineering implementation, ensuring high code quality and performance. Salary range is between $500,000 to $850,000 annually.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Code RL Research Engineer — Build Safe, Scalable AI
Code RL Research Engineer — Build Safe, Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Code RL Research Engineer — AI Coding & RL Systems
Code RL Research Engineer — AI Coding & RL Systems

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Code RL Research Engineer — End-to-End Coding Pioneer
Code RL Research Engineer — End-to-End Coding Pioneer

Neura Market • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer, Code RL: Build & Validate AI Code
Research Engineer, Code RL: Build & Validate AI Code

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Performance RL Research Engineer - Safe Code & AI (Hybrid)
Performance RL Research Engineer - Safe Code & AI (Hybrid)

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Code RL Research Engineer: Build Safe, Fast AI Code
Code RL Research Engineer: Build Safe, Fast AI Code

Jobzhr • San Francisco (CA), Northern (KY)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning) Anthropic San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) Anthropic San Francisco, CA | New York City, NY

Neura Market • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2