Founding Reinforcement Learning Engineer

Clera

San Francisco (CA)

Presencial

USD 125.000 - 200.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca para este puesto — genera un currículum y una carta de presentación adaptados en cuestión de un minuto.

Supera los filtros ATS

Descripción de la vacante

Clera in San Francisco, CA is building core reinforcement learning systems as a founding engineer on an early-stage AI team. You will work directly with the founders and own the process from environment design and model training through evaluation.

You design RL environments, train models with PPO, GRPO, DPO, RLHF, and RLAIF, and develop evaluation frameworks. You will run experiments, interpret results, and decide what approaches to explore next, scaling training on GPU clusters.

Formación

  • 2+ years hands-on RL or ML engineering (could range 2–10+ years).
  • Strong Python and deep experience with PyTorch or JAX.
  • Experience training models with reinforcement learning (policy gradients, reward modeling, RLHF).

Responsabilidades

  • Design and build RL environments, reward functions, and training pipelines.
  • Train and fine-tune models using PPO, GRPO, DPO, RLHF, and RLAIF.
  • Develop evaluation frameworks to measure model and agent performance.
  • Run experiments, interpret results, and decide next exploration steps.
  • Scale training on GPU clusters and maintain reliable pipelines.
  • Turn research ideas into production systems.
  • Help establish engineering culture and hire future engineers.

Conocimientos

Python
Reinforcement learning
Distributed training

Educación

Bachelor in CS / Math / Physics

Herramientas

PyTorch
JAX
Ray
CUDA
Kubernetes
Gymnasium
RLlib
MuJoCo
TRL
verl
OpenRLHF

Descripción del empleo

About the Role

Build core reinforcement learning systems from the ground up as a founding engineer on an early-stage AI team. You will work directly with the founders and own the process from environment design and model training through evaluation, helping shape the company's technical direction.

What You'll Do
  • Design and build reinforcement learning environments, reward functions, and training pipelines.
  • Train and fine-tune models using methods such as PPO, GRPO, DPO, RLHF, and RLAIF.
  • Develop evaluation frameworks to measure model and agent performance.
  • Run experiments, interpret results, and decide which approaches to explore next.
  • Scale training on GPU clusters and maintain reliable pipelines.
  • Turn research ideas into production systems.
  • Help establish engineering culture and hire future engineers.
What We're Looking For
  • At least 2 years of hands‑on experience in reinforcement learning or machine learning engineering, with relevant experience potentially ranging from 2 to 10 or more years.
  • Strong Python skills and deep experience with PyTorch or JAX.
  • Practical experience training models with reinforcement learning, including policy gradient methods, reward modeling, or RLHF.
  • Comfort with distributed training and GPU infrastructure, including tools such as Ray, CUDA, and Kubernetes.
  • Experience with LLM post‑training or agent training is a plus, as is familiarity with Gymnasium, Ray RLlib, Isaac, MuJoCo, TRL, verl, or OpenRLHF.
  • A degree in computer science, mathematics, physics, or a related field is sought; a master's or PhD is a plus. Publications or open‑source work in RL are also valued.
  • A hands‑on builder who enjoys moving quickly in a small, early‑stage team.
Compensation & Benefits

Compensation is $125,000 to $200,000 USD annually.

Location

On-site in San Francisco, California, United States.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Founding Reinforcement Learning Engineer
Founding Reinforcement Learning Engineer

Wwshemi • San Francisco (CA)

Presencial
USD 125.000 - 200.000
Founding Reinforcement Learning Engineer Clera · San Francisco, CA Full-time · On-site $125,000–200,000 3 hours ago
Founding Reinforcement Learning Engineer Clera · San Francisco, CA Full-time · On-site $125,000–200,000 3 hours ago

Emploive • San Francisco (CA)

Presencial
USD 125.000 - 200.000
Founding RL Engineer
Founding RL Engineer

Clera Labs, Inc. • San Francisco (CA)

Presencial
USD 125.000 - 200.000
Founding RL Engineer – Build Core RL Systems (SF)
Founding RL Engineer – Build Core RL Systems (SF)

Clera Labs, Inc. • San Francisco (CA)

Presencial
USD 125.000 - 200.000
Software Engineer, RL Environments
Software Engineer, RL Environments

Wintermeyer Ventures • San Francisco (CA)

Presencial
USD 260.000 - 290.000
Full-Stack Software Engineer, Reinforcement Learning (Mid-Level)
Full-Stack Software Engineer, Reinforcement Learning (Mid-Level)

Clera • San Francisco (CA)

Presencial
USD 130.000 - 225.000
Health coverage
Paid time off
Retirement benefits
+1
Machine Learning Engineer — Reinforcement Learning
Machine Learning Engineer — Reinforcement Learning

Lever, Inc. • Sunnyvale (CA)

Presencial
USD 150.000 - 450.000
Medical benefits
Dental benefits
Vision benefits
+7
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

Presencial
USD 180.000 - 220.000
RL Environment Software Engineer
RL Environment Software Engineer

TalentPluto, Inc. • San Francisco (CA)

A distancia
USD 180.000 - 220.000
Research Engineer, Training and Environment Infrastructure
Research Engineer, Training and Environment Infrastructure

Ersilia • San Francisco (CA)

Presencial
USD 150.000 - 350.000
Meaningful equity grants
Health, dental, and vision coverage