Reinforcement Learning Research Engineer — Safe AI Code

SignalAI

San Francisco (CA)

Hybrid

USD 350,000 - 850,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity donation matching
Generous vacation & parental leave
Flexible working hours
Office in San Francisco

Job summary

Anthropic is hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators.

You’ll Need To Know Accelerator Performance Well To Turn It Into Tasks And Signals Models Can Learn From. Specifically, You Will Invent, design and implement RL environments and evaluations, conduct experiments and shape our research roadmap, deliver your work into training runs, and collaborate with other

Qualifications

  • Experience with accelerators (CUDA, ROCm, Triton, Pallas).
  • Experience with ML frameworks (JAX or PyTorch).
  • Experience across kernels, model code, and distributed systems.

Responsibilities

  • Invent, design and implement RL environments and evaluations.
  • Conduct experiments and shape our research roadmap.
  • Deliver your work into training runs.
  • Collaborate with other researchers, engineers, and performance engineering specialists across and outside Anthropic.

Skills

Accelerators
ML frameworks
Code & kernel stacks
Distributed systems
Research to engineering balance

Education

Bachelor’s degree

Job description

Anthropic is hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators.

You’ll Need To Know Accelerator Performance Well To Turn It Into Tasks And Signals Models Can Learn From. Specifically, You Will Invent, design and implement RL environments and evaluations, conduct experiments and shape our research roadmap, deliver your work into training runs, and collaborate with other

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Performance RL Research Engineer - Safe Code for Accelerator
Performance RL Research Engineer - Safe Code for Accelerator

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Competitive compensation
Flexible hours
Code RL Research Engineer: Build Safe, Fast AI Code
Code RL Research Engineer: Build Safe, Fast AI Code

Jobzhr • San Francisco (CA), Northern (KY)

Hybrid
USD 500,000 - 850,000
Performance RL Research Engineer: Safe Code for Accelerators
Performance RL Research Engineer: Safe Code for Accelerators

Menlo Ventures • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Performance RL Research Engineer - Safe Code & AI (Hybrid)
Performance RL Research Engineer - Safe Code & AI (Hybrid)

Anthropic • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning)
Research Engineer, Code RL (Reinforcement Learning)

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Research Engineer, Performance RL (Reinforcement Learning)
Research Engineer, Performance RL (Reinforcement Learning)

Menlo Ventures • San Francisco (CA)

Hybrid
USD 350,000 - 850,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Research Engineer, Performance RL (Reinforcement Learning) Anthropic San Francisco, CA
Research Engineer, Performance RL (Reinforcement Learning) Anthropic San Francisco, CA

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Competitive compensation
Flexible hours
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Code RL Research Engineer — Build Safe, Scalable AI
Code RL Research Engineer — Build Safe, Scalable AI

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000