Production RL Engineer: Train & Deploy Web-Data Models

Firecrawl

San Francisco (CA)

Hybrid

USD 180,000 - 270,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Up to 0.15% equity
Generous PTO
Parental leave
Wellness stipend
Learning & Development budget

Job summary

A leading data extraction company is seeking a Research Engineer focused on reinforcement learning in San Francisco or Remote. In this full-time role, you will build training infrastructures, fine-tune models, and bridge classical RL and modern agent systems. Ideal candidates have 3+ years in applied RL or ML engineering and are comfortable communicating across teams. Competitive salary of $180,000–$270,000 with great benefits including equity, generous PTO, and wellness stipends.

Qualifications

  • 3+ years of experience in applied reinforcement learning, ML engineering, or model training with production systems.
  • Experience operating GPU clusters, managing training runs, and debugging convergence issues.

Responsibilities

  • Build training infrastructure and reward pipelines from scratch.
  • Fine-tune models to achieve state-of-the-art results.
  • Run fast experiments and iterate quickly.
  • Communicate RL concepts clearly to non-RL stakeholders.
  • Collaborate with the research team.

Skills

Reinforcement learning
Machine learning engineering
Model training
Data pipelines
Experiment design

Job description

A leading data extraction company is seeking a Research Engineer focused on reinforcement learning in San Francisco or Remote. In this full-time role, you will build training infrastructures, fine-tune models, and bridge classical RL and modern agent systems. Ideal candidates have 3+ years in applied RL or ML engineering and are comfortable communicating across teams. Competitive salary of $180,000–$270,000 with great benefits including equity, generous PTO, and wellness stipends.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer - Post-Training & Data Environments
Research Engineer - Post-Training & Data Environments

Mercor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Generous equity grant
$10K housing bonus
$1.5K monthly food stipend
+2
Infrastructure Research Engineer - Large-Scale RL Systems
Infrastructure Research Engineer - Large-Scale RL Systems

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research Software Engineer — Scalable RL & Distributed Training
Research Software Engineer — Scalable RL & Distributed Training

Reflection AI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Research Engineer: Product-Driven ML & RL
Research Engineer: Product-Driven ML & RL

OpenAI • San Francisco (CA)

Hybrid
USD 295,000 - 555,000
Software Engineer, RL Environments & Evaluation Pipelines
Software Engineer, RL Environments & Evaluation Pipelines

davidjoseph-co • San Francisco (CA)

On-site
USD 180,000 - 220,000
Competitive equity
Principal RL Environment Engineer
Principal RL Environment Engineer

Turing • San Francisco (CA)

On-site
USD 250,000 - 350,000
Competitive compensation
Collaborative work culture
Equity options
Software Engineer — RL Environments for Frontier AI
Software Engineer — RL Environments for Frontier AI

Mechanize, Inc. • San Francisco (CA)

On-site
USD 100,000 - 140,000
401k
Health insurance
Dental insurance
+2
Reinforcement Learning Data & Environments Engineer
Reinforcement Learning Data & Environments Engineer

Mirendil • San Francisco (CA)

On-site
USD 350,000 - 500,000
Research Engineer, Code RL: Build & Validate AI Code
Research Engineer, Code RL: Build & Validate AI Code

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 500,000 - 850,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

Hybrid
USD 180,000 - 220,000