Staff Software Engineer - RL Frameworks & Tooling

Anthropic Limited

New York (NY)

Hybrid

USD 405,000 - 625,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Anthropic is seeking an experienced engineer to shape reinforcement learning tooling and APIs for agentic coding environments. You will embed with research teams, design frameworks, and maintain production RL runs. Strong Python, API design, and the ability to work in large codebases are essential.

Location-based hybrid policy requires in-office presence at least 25% of the time. You’ll collaborate across engineers and researchers, translating research needs into robust, scalable software that

Qualifications

  • Bachelor’s degree in a field relevant to the role.
  • Strong Python expertise and ability to design APIs.
  • Experience with research-style or large codebases.

Responsibilities

  • Design widely-used APIs, frameworks, and abstractions for engineers and researchers.
  • Embed with research teams on a rotational basis and transfer ownership.
  • Work directly in research codebases to improve reliability and structure.
  • Anticipate silent failure modes and prevent them structurally via typing and testing.
  • Contribute to reliability and triage tooling for production RL systems.
  • Help define engineering standards, review practices, and design patterns for a new team.
  • Demonstrate deep Python expertise, including typing and async patterns.
  • Show a track record designing intuitive, safe APIs or frameworks.
  • Experience working in large, evolving codebases you didn’t write initially.
  • Anticipate failure modes and prevent them via system design and testing.
  • Strong written and verbal communication skills for collaboration.
  • Comfort with ambiguity and driving outcomes from loosely defined problems.

Skills

Python
API design
Framework design

Education

Bachelor's degree

Job description

Anthropic is seeking an experienced engineer to shape reinforcement learning tooling and APIs for agentic coding environments. You will embed with research teams, design frameworks, and maintain production RL runs. Strong Python, API design, and the ability to work in large codebases are essential.

Location-based hybrid policy requires in-office presence at least 25% of the time. You’ll collaborate across engineers and researchers, translating research needs into robust, scalable software that

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer — RL Infrastructure & Platforms
Staff Software Engineer — RL Infrastructure & Platforms

Anthropic • New York (NY), Seattle (WA), San Francisco (CA)

On-site
USD 140,000 - 180,000
Staff Software Engineer — RL Environment Platform
Staff Software Engineer — RL Environment Platform

Anthropic • New York (NY)

Hybrid
USD 405,000 - 605,000
Staff Software Engineer, Code RL
Staff Software Engineer, Code RL

Anthropic • New York (NY), Seattle (WA), San Francisco (CA)

On-site
USD 140,000 - 180,000
Staff Software Engineer, Environments Infra — Scale RL
Staff Software Engineer, Environments Infra — Scale RL

Anthropic Limited • San Francisco (CA)

Hybrid
USD 405,000 - 605,000
Staff AI RL Infrastructure Engineer
Staff AI RL Infrastructure Engineer

Anthropic • San Francisco (CA)

Hybrid
USD 405,000 - 605,000
Competitive compensation
Equity donation matching (optional)
Vacation and parental leave
+2
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer, Code RL (Reinforcement Learning) Anthropic San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) Anthropic San Francisco, CA | New York City, NY

Neura Market • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Code RL Research Engineer — AI Coding & RL Systems
Code RL Research Engineer — AI Coding & RL Systems

Anthropic • New York (NY)

Hybrid
USD 500,000 - 850,000
Code RL Research Engineer: Build Safe, Fast AI Code
Code RL Research Engineer: Build Safe, Fast AI Code

Jobzhr • San Francisco (CA), Northern (KY)

Hybrid
USD 500,000 - 850,000
Research Engineer, Performance RL (Reinforcement Learning) Anthropic San Francisco, CA
Research Engineer, Performance RL (Reinforcement Learning) Anthropic San Francisco, CA

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 230,000
Competitive compensation
Flexible hours