Full Stack Software Engineer - RL environments

Acceler8 Talent

San Francisco (CA)

On-site

USD 120,000 - 180,000

Full time

31 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Acceler8 Talent is seeking a Software Engineer focused on RL Environments to join a fast-growing applied AI research company in San Francisco. You’ll help design the environments, reward signals, evaluations, and data that influence how frontier models are trained and improved.

You will build RL environments, simulations, and task frameworks, develop reward signals and evaluation systems for RLHF / RLVR, and work with researchers to turn training objectives into production systems using Docker

Qualifications

  • Strong reinforcement learning experience and hands-on work building RL environments, reward signals, and model evaluation methods.
  • Experience delivering RL production systems and data pipelines in fast-moving teams.
  • Familiarity with Docker, Kubernetes, and scalable infra for running experiments at scale.
  • Experience in RL environments, AI evaluation, benchmarking, or AI safety oriented organizations is a plus.
  • High ownership mindset; can thrive in a startup-like setting with rapid iteration.

Responsibilities

  • Build RL environments, simulations, and task frameworks
  • Develop reward signals and evaluation systems for RLHF / RLVR
  • Analyse model and agent failure modes
  • Build synthetic and real-world data pipelines
  • Work closely with researchers to turn training objectives into production systems

Skills

Reinforcement learning
Environments
Rewards
Model evaluations
Production systems
RLHF / RLVR
Startup experience
AI research

Tools

Docker
Kubernetes
Python
ML frameworks

Job description

Software Engineer — RL Environments | San Francisco

We’re hiring a Software Engineer focused on RL Environments to join a fast-growing applied AI research company working directly with leading frontier model labs.

You’ll help design the environments, reward signals, evaluations, and data that influence how advanced models are trained and improved.

What you’ll do
  • Build RL environments, simulations, and task frameworks
  • Develop reward signals and evaluation systems for RLHF / RLVR
  • Analyse model and agent failure modes
  • Build synthetic and real-world data pipelines
  • Work closely with researchers to turn training objectives into production systems
What we’re looking for
  • Strong reinforcement learning experience
  • Experience building environments, rewards, or model evaluations
  • Ability to build and deploy production systems
  • Experience with Docker, Kubernetes, or similar infrastructure
  • Fast-moving, high-ownership mindset
  • Experience at an RL environment, AI evaluation, benchmarking, or AI safety organisation
  • High-growth startup, early engineer, or founder experience
  • Exceptional technical achievement or standout internships

Primarily targeting candidates with 1–5 years of experience, although exceptional new graduates may be considered.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

On-site
USD 180,000 - 220,000
Software Engineer
Software Engineer

Acceler8 Talent • San Francisco (CA)

On-site
USD 225,000 - 275,000
Equity
Reinforcement Learning Environments Engineer
Reinforcement Learning Environments Engineer

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 180,000
SWE (RL Environments) "Reinforcement Learning"
SWE (RL Environments) "Reinforcement Learning"

AI Talent Now • San Francisco (CA)

On-site
USD 150,000 - 250,000
Research Engineer – RL Infrastructure & Agent Environments
Research Engineer – RL Infrastructure & Agent Environments

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 150,000 - 190,000
RL Environment Engineer: Shape Frontier AI
RL Environment Engineer: Shape Frontier AI

AI Talent Now • San Francisco (CA)

On-site
USD 150,000 - 250,000
RL Environments Engineer
RL Environments Engineer

Bespoke Labs Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 250,000 - 300,000
Health, dental, and vision coverage
401(k)
Daily onsite lunch provided
+2
RL Environments Engineer
RL Environments Engineer

Bespoke-Labs • Mountain View (CA)

On-site
USD 250,000 - 300,000
Health, dental, and vision
401(k)
Daily onsite lunch
+3
Research Engineer, RL Environments and Infrastructure
Research Engineer, RL Environments and Infrastructure

Hyphen Connect • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Staff Software Engineer, RL Environments
Staff Software Engineer, RL Environments

United States Digital Space LLC • San Francisco (CA), New York (NY)

On-site
USD 252,000 - 315,000
Health insurance
Dental & vision coverage
:Learning stipend