RL Research Engineer - Scalable, Safe AI Systems

Anthropic

San Francisco (CA)

Hybrid

USD 500,000 - 850,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Equity donation matching
Generous vacation and parental leave
Flexible working hours
Office space

Job summary

Anthropic in San Francisco seeks a Research Engineer in Reinforcement Learning to advance the capabilities and safety of large language models, blending research and engineering. You will implement novel approaches, build scalable infrastructure, and develop prototype tools for internal use, productivity, and evaluation.

Strong candidates have Python, ML frameworks, and experience in ML research. The role offers a hybrid work arrangement with a focus on collaboration and impact.

Qualifications

  • Proficient in Python and async/concurrent programming with frameworks like Trio.
  • Experience with machine learning frameworks (PyTorch, TensorFlow, JAX).
  • Industry experience in machine learning research.

Responsibilities

  • Architect and optimize core reinforcement learning infrastructure and experiment management.
  • Design and test new training environments and evaluation methods for RL agents.
  • Collaborate across research and engineering to build scalable AI research infrastructure.

Skills

Python
Async programming
Pair programming
Communication skills

Education

Bachelor's degree

Tools

PyTorch
TensorFlow
JAX
Kubernetes
Rust
C++
Virtualization

Job description

Anthropic in San Francisco seeks a Research Engineer in Reinforcement Learning to advance the capabilities and safety of large language models, blending research and engineering. You will implement novel approaches, build scalable infrastructure, and develop prototype tools for internal use, productivity, and evaluation.

Strong candidates have Python, ML frameworks, and experience in ML research. The role offers a hybrid work arrangement with a focus on collaboration and impact.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer: ML & RL for Safe, Scalable AI
Research Engineer: ML & RL for Safe, Scalable AI

Anthropic • New York (NY)

On-site
USD 280,000 - 425,000
Research Engineer - Large-Scale ML & Safe AI
Research Engineer - Large-Scale ML & Safe AI

Anthropic • New York (NY)

Hybrid
USD 350,000 - 500,000
Code RL Research Engineer — Build Safe, Scalable AI
Code RL Research Engineer — Build Safe, Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer: Product‑Driven ML for Safer AI
Research Engineer: Product‑Driven ML for Safer AI

OpenAI • California (MO)

Hybrid
USD 180,000 - 240,000
Relocation assistance
Hybrid work model
Research Engineer for Safe, Scalable LLMs
Research Engineer for Safe, Scalable LLMs

Anthropic • Seattle (WA)

On-site
USD 350,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation
+3
LLM Research Scientist - Scaling AI & Safety
LLM Research Scientist - Scaling AI & Safety

AI Breaking Wire • Menlo Park (CA)

On-site
USD 180,000 - 300,000
Competitive salary
Equity
Health benefits
+1
Research Scientist - AI Safety & Scalable ML
Research Scientist - AI Safety & Scalable ML

Center for AI Safety (CAIS) • San Francisco (CA)

On-site
USD 140,000 - 200,000
Health insurance for you and dependets
401K plan + 4% matching
Unlimited PTO
+2
Code RL Research Engineer — AI Coding & RL Systems
Code RL Research Engineer — AI Coding & RL Systems

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
ML Systems Engineer — RL Training & Finetuning
ML Systems Engineer — RL Training & Finetuning

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Staff Research Engineer — AI Discovery & Scalable Systems
Staff Research Engineer — AI Discovery & Scalable Systems

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000