RL Researcher, Product Systems

Autohand AI Ltd.

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Publications

Job summary

Autohand AI Ltd. in San Francisco is seeking a Reinforcement Learning Researcher with a strong product mindset to join our research team.

You'll develop RL systems that enable autonomous programming agents to learn and improve from real-world feedback. The role emphasizes designing reward functions, training pipelines, and evaluation frameworks that push the boundaries of what autonomous systems can achieve.

Qualifications

  • PhD or equivalent in RL/ML is required.
  • Strong track record in reinforcement learning research.
  • Experience shipping ML systems to production.
  • Product-oriented mindset with ability to prioritize impactful research.

Responsibilities

  • Design and implement RL algorithms for autonomous code generation and modification.
  • Develop reward models and training pipelines that align with product goals.
  • Build evaluation frameworks to measure agent performance and safety.
  • Collaborate with product teams to understand user needs and translate them into research directions.
  • Ship research to production and iterate based on real-world feedback.
  • Contribute to technical publications and open-source releases.

Skills

Product-oriented mindset
Strong RL research track record
Excellent communication
Cross-functional collaboration

Education

PhD or equivalent in RL/ML

Tools

Python
PyTorch
JAX

Job description

We're looking for a Reinforcement Learning Researcher with a strong product mindset to join our research team. You'll develop RL systems that enable autonomous programming agents to learn and improve from real-world feedback.

This role is ideal for someone who wants their research to directly impact products. You'll design reward functions, training pipelines, and evaluation frameworks that push the boundaries of what autonomous systems can achieve.

What you'll do
  • Design and implement RL algorithms for autonomous code generation and modification
  • Develop reward models and training pipelines that align with product goals
  • Build evaluation frameworks to measure agent performance and safety
  • Collaborate with product teams to understand user needs and translate them into research directions
  • Ship research to production and iterate based on real-world feedback
  • Contribute to technical publications and open-source releases
What we're looking for
  • PhD or equivalent experience in RL, ML, or related fields
  • Strong track record in reinforcement learning research
  • Experience shipping ML systems to production
  • Product-oriented mindset with ability to prioritize impactful research
  • Proficiency in Python and deep learning frameworks (PyTorch, JAX)
  • Excellent communication skills and ability to work cross-functionally
Nice to have
  • Publications in top ML venues (NeurIPS, ICML, ICLR, CoRL)
  • Experience with RLHF or preference learning
  • Background in LLMs or code generation
  • Experience with distributed training at scale
  • Previous startup or product-focused research experience
Preferred Candidates

You must be a citizen or have authorised work permit in one of the locations where we hire: Australia, New Zealand, Brazil, Singapore, or the United States.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist - Reinforcement Learning
Research Scientist - Reinforcement Learning

Optimized, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 185,000 - 255,000
Research Engineer / Scientist – Reinforcement Learning (RL)
Research Engineer / Scientist – Reinforcement Learning (RL)

Percepta • New York (NY)

On-site
USD 110,000 - 150,000
Research Engineer
Research Engineer

Bespoke Labs • Mountain View (CA)

On-site
USD 120,000 - 140,000
Health coverage
Opportunity to work with leading AI labs
Competitive salary and equity
Research Scientist, Reinforcement Learning
Research Scientist, Reinforcement Learning

DeepRoute • Fremont (CA)

On-site
USD 90,000 - 130,000
Research Scientist, Reinforcement Learning
Research Scientist, Reinforcement Learning

Deeproute.ai • Colorado

On-site
USD 100,000 - 150,000
Applied RL Researcher for Production Systems
Applied RL Researcher for Production Systems

Autohand AI Ltd. • San Francisco (CA)

On-site
USD 150,000 - 210,000
Publications
AIML - Machine Learning Research Lead, RL Agents, MLR
AIML - Machine Learning Research Lead, RL Agents, MLR

Socket.dev • Cupertino (CA)

On-site
USD 220,000 - 320,000
Research Scientist, LLM Evaluation & Post-Training
Research Scientist, LLM Evaluation & Post-Training

OneForma • United States

Hybrid
USD 140,000 - 210,000
Staff Reinforcement Learning Engineer
Staff Reinforcement Learning Engineer

Fruition Group US • California (MO)

On-site
USD 150,000 - 210,000
Machine Learning Researcher
Machine Learning Researcher

Brahma Consulting Group • San Francisco (CA)

On-site
USD 150,000 - 230,000