Code RL Research Engineer — AI Coding & RL Systems

Anthropic

New York (NY)

Hybrid

USD 500,000 - 850,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Anthropic is hiring a Research Engineer for the Code RL team in New York. Applicants should possess strong software engineering skills and deep expertise in Python to design RL environments and run training experiments. The position requires balancing research and engineering responsibilities while ensuring code quality and performance.

With an annual salary ranging from $500,000 to $850,000, Anthropic aims to build reliable and beneficial AI systems. Candidates must have at least a Bachelor's degree or equivalent experience.

Qualifications

  • Have strong software engineering skills and deep Python expertise, including async and concurrent programming.
  • Balance research exploration with engineering implementation, and rigorously design experiments and interpret results.
  • Experience with reinforcement learning, RLHF, or large-scale distributed training.

Responsibilities

  • Design RL environments and coding tasks, build reward signals and verifiers.
  • Run training experiments on frontier models and diagnose model behavior.
  • Enhance speed and reliability of training pipelines.

Skills

Strong software engineering skills
Deep Python expertise
Research exploration and engineering implementation
Code quality and performance awareness

Education

Bachelor’s degree or equivalent

Tools

PyTorch
CUDA / GPU or TPU

Job description

Anthropic is hiring a Research Engineer for the Code RL team in New York. Applicants should possess strong software engineering skills and deep expertise in Python to design RL environments and run training experiments. The position requires balancing research and engineering responsibilities while ensuring code quality and performance.

With an annual salary ranging from $500,000 to $850,000, Anthropic aims to build reliable and beneficial AI systems. Candidates must have at least a Bachelor's degree or equivalent experience.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Code RL Research Engineer — Build Safe, Scalable AI
Code RL Research Engineer — Build Safe, Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

On-site
USD 500,000 - 850,000
Research Engineer, Machine Learning (RL Velocity) Anthropic · Remote · US · AI Research $500,000–$850,000 5mo ago
Research Engineer, Machine Learning (RL Velocity) Anthropic · Remote · US · AI Research $500,000–$850,000 5mo ago

Aimlroles • Northern (KY)

Hybrid
USD 500,000 - 850,000
Research Scientist, RL for Coding Agents
Research Scientist, RL for Coding Agents

cursor • New York (NY), San Francisco (CA)

On-site
USD 100,000 - 130,000
Research Engineer, Chip Design RL (Reinforcement Learning)
Research Engineer, Chip Design RL (Reinforcement Learning)

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Flexible hours
Generous vacation
Parental leave
+2
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • Seattle (WA)

On-site
USD 520,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • New York (NY)

On-site
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer (Cybersecurity Reinforcement Learning)
Research Engineer (Cybersecurity Reinforcement Learning)

SupportFinity™ • San Francisco (CA)

On-site
USD 300,000 - 405,000
Competitive compensation and benefits.
Optional equity donation matching.
Generous vacation and parental leave.
+2