Research Engineer, Reinforcement Learning

Techire Ai

San Francisco (CA)

On-site

USD 270,000 - 330,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology solutions provider in San Francisco seeks an experienced professional to design and build large-scale reinforcement learning environments. The role involves collaborating with researchers to tackle complex problems, refining learning processes, and scaling infrastructure. Ideal candidates will have extensive machine learning and reinforcement learning experience, a strong intuition for agent dynamics, and a desire to explore challenging scenarios. The compensation package is competitive, with a base salary of up to $300K plus equity.

Qualifications

  • Strong ML and RL experience is essential.
  • Deep understanding of agent dynamics is required.
  • Curiosity and problem-solving skills are needed.

Responsibilities

  • Build reinforcement learning environments for training agents.
  • Collaborate with researchers on unsolved challenges.
  • Scale infrastructure and refine learning processes.

Skills

Machine Learning (ML)
Reinforcement Learning (RL)
Intuition for agent dynamics
Curiosity to explore complex problems

Job description

Want to build the large-scale RL environments frontier labs use to train agents that can truly reason and act?

This team are creating complex reinforcement learning environments — simulations where advanced agents learn to plan, adapt, and solve multi-step problems that stretch beyond standard benchmarks. The focus isn’t on training the models themselves, but on building the worlds that make meaningful learning and evaluation possible — the foundation for more capable, aligned systems.

You’ll work end-to-end across environment design, reward dynamics, and scalable simulation — developing the feedback loops that define what “good” looks like for intelligent behaviour. It’s open-ended, research-driven work where the task definition, data, and reward structure are often the hardest and most important problems to solve.

You’ll collaborate closely with researchers tackling unsolved challenges in reinforcement learning and agent behaviour, shaping experiments, scaling infrastructure, and refining how agents learn in the loop.

It suits someone with strong ML and RL experience, deep intuition for agent dynamics, and the curiosity to explore problems that don’t come with clear instructions.

On-site in San Francisco. Compensation up to $300 K base (negotiable, depending on experience) plus equity.

If you want to help build the environments that teach the next generation of AI systems how to think, act, and adapt — we’d love to hear from you.

All applicants will receive a response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

Hybrid
USD 180,000 - 220,000
SWE (RL Environments) "Reinforcement Learning"
SWE (RL Environments) "Reinforcement Learning"

AI Talent Now • San Francisco (CA)

On-site
USD 150,000 - 250,000
Research Engineer
Research Engineer

Barrington James • San Francisco (CA)

On-site
USD 140,000 - 210,000
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Research Engineer/Research Scientist, RL
Research Engineer/Research Scientist, RL

OpenAI • San Francisco (CA)

Hybrid
USD 310,000 - 460,000
Medical, dental, and vision insurance
Mental health and wellness support
401(k) plan with 50% matching
+3
Researcher, Synthetic RL
Researcher, Synthetic RL

OpenAI • Los Angeles (CA)

Hybrid
USD 120,000 - 160,000
RL Environment Engineer: Shape Frontier AI
RL Environment Engineer: Shape Frontier AI

AI Talent Now • San Francisco (CA)

On-site
USD 150,000 - 250,000
Research Engineer, Infrastructure, RL Systems
Research Engineer, Infrastructure, RL Systems

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research Engineer (Mountain View) - 17813
Research Engineer (Mountain View) - 17813

somewhere • Mountain View (CA)

On-site
USD 120,000 - 160,000
Health coverage
Ownership upside
Collaboration with leading AI research organizations
Research Engineer
Research Engineer

Bespoke Labs • Mountain View (CA)

On-site
USD 120,000 - 140,000
Health coverage
Opportunity to work with leading AI labs
Competitive salary and equity