Founding Reinforcement Learning Engineer Clera · San Francisco, CA Full-time · On-site $125,000–200,000 3 hours ago

Emploive

San Francisco (CA)

On-site

USD 125,000 - 200,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Emploive is building core reinforcement learning systems and seeks a founding engineer to own the end-to-end development from environments to production. You will work directly with the founders, shape technical direction, and move quickly in a small, early-stage team.

The role emphasizes hands-on building, distributed training, and collaboration with a tight-knit team to evolve the platform and hire future engineers.

Qualifications

  • 2+ years hands-on experience in reinforcement learning or ML engineering.
  • Experience with policy gradient methods, reward modeling, or RLHF preferred.
  • Strong Python skills and experience with PyTorch or JAX required.

Responsibilities

  • Design and build reinforcement learning environments, reward functions, and training pipelines.
  • Train and fine-tune models using PPO, GRPO, DPO, RLHF, and RLAIF.
  • Develop evaluation frameworks to measure model and agent performance.
  • Run experiments, interpret results, and decide which approaches to explore next.
  • Scale training on GPU clusters and maintain reliable pipelines.
  • Turn research ideas into production systems.
  • Help establish engineering culture and hire future engineers.

Skills

Python
PyTorch
JAX
Reinforcement learning
Distributed training
Ray
CUDA
Kubernetes

Education

Bachelor's degree in CS/Math/Physics
Master's or PhD is a plus

Tools

Gymnasium
Ray RLlib
MuJoCo
OpenRLHF
TRL

Job description

About the Role

Build core reinforcement learning systems from the ground up as a founding engineer on an early-stage AI team. You will work directly with the founders and own the process from environment design and model training through evaluation, helping shape the company's technical direction.


What You'll Do


  • Design and build reinforcement learning environments, reward functions, and training pipelines.

  • Train and fine-tune models using methods such as PPO, GRPO, DPO, RLHF, and RLAIF.

  • Develop evaluation frameworks to measure model and agent performance.

  • Run experiments, interpret results, and decide which approaches to explore next.

  • Scale training on GPU clusters and maintain reliable pipelines.

  • Turn research ideas into production systems.

  • Help establish engineering culture and hire future engineers.


What We're Looking For


  • At least 2 years of hands-on experience in reinforcement learning or machine learning engineering, with relevant experience potentially ranging from 2 to 10 or more years.

  • Strong Python skills and deep experience with PyTorch or JAX.

  • Practical experience training models with reinforcement learning, including policy gradient methods, reward modeling, or RLHF.

  • Comfort with distributed training and GPU infrastructure, including tools such as Ray, CUDA, and Kubernetes.

  • Experience with LLM post-training or agent training is a plus, as is familiarity with Gymnasium, Ray RLlib, Isaac, MuJoCo, TRL, verl, or OpenRLHF.

  • A degree in computer science, mathematics, physics, or a related field is sought; a master's or PhD is a plus. Publications or open-source work in RL are also valued.

  • A hands‑on builder who enjoys moving quickly in a small, early‑stage team.


Compensation & Benefits

Compensation is $125,000 to $200,000 USD annually.


Location

On-site in San Francisco, California, United States.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Founding Reinforcement Learning Engineer
Founding Reinforcement Learning Engineer

Clera • San Francisco (CA)

On-site
USD 125,000 - 200,000
Founding Reinforcement Learning Engineer
Founding Reinforcement Learning Engineer

Wwshemi • San Francisco (CA)

On-site
USD 125,000 - 200,000
Founding RL Engineer: Shape Core AI Systems
Founding RL Engineer: Shape Core AI Systems

Clera • San Francisco (CA)

On-site
USD 125,000 - 200,000
Full-Stack Software Engineer, Reinforcement Learning (Mid-Level)
Full-Stack Software Engineer, Reinforcement Learning (Mid-Level)

Clera • San Francisco (CA)

On-site
USD 130,000 - 225,000
Health coverage
Paid time off
Retirement benefits
+1
Software Engineer, RL Environments
Software Engineer, RL Environments

Wintermeyer Ventures • San Francisco (CA)

On-site
USD 260,000 - 290,000
Senior Applied Reinforcement Learning Engineer
Senior Applied Reinforcement Learning Engineer

Centific • Palo Alto (CA), Northern (KY)

Hybrid
USD 150,000 - 300,000
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY
Research Engineer, Code RL (Reinforcement Learning) San Francisco, CA | New York City, NY

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

On-site
USD 180,000 - 220,000
RL Environment Software Engineer
RL Environment Software Engineer

TalentPluto, Inc. • San Francisco (CA)

Remote
USD 180,000 - 220,000
Senior Applied Reinforcement Learning Engineer
Senior Applied Reinforcement Learning Engineer

Centific Global Solutions, Inc. • United States

Hybrid
USD 150,000 - 300,000