RL Research Engineer: Scalable Training & Eval (Remote)

24-MAG

United States

Remote

USD 400,000 - 800,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

24-MAG LLC is seeking a remote Member of Technical Staff in Research Engineering with deep RL expertise. The role spans RL environment design, scalable training pipelines, synthetic data generation, automated evaluation, and benchmarking to accelerate AI research and development.

You will work at the interface of research and production, building RL environments, training workflows, and evaluation systems that improve model capability, reliability, and research velocity in a fully remote,

Qualifications

  • Deep professional or research experience in reinforcement learning.
  • Strong understanding of RL environment design, reward structures, training dynamics, and evaluation.
  • Demonstrated experience building and scaling RL systems, training pipelines, or experimentation frameworks.
  • Strong experience with automation and synthetic data-generation workflows.
  • Familiarity with automated evaluation, model validation, and quality-assurance systems.
  • Experience fine-tuning and evaluating open-source machine-learning models.
  • Strong technical writing and communication skills.
  • Ability to operate effectively in fast-paced, research-driven environments.

Responsibilities

  • Reinforcement Learning Environment Design: architect self-contained RL environments and design rewards, verifiers, and evaluation logic.
  • Training Pipelines & Experimentation Systems: design scalable episode pipelines and reproducible experimentation workflows.
  • Synthetic Data & Automated Evaluation: build automated data-generation systems and automated evaluation and QA systems.
  • Model Optimisation & Benchmarking: fine-tune open-source RL/ML models and develop benchmarking frameworks.
  • Cross-functional collaboration to translate research goals into robust, production-ready systems.

Skills

Reinforcement Learning
RL Environments
Training Pipelines
Experimentation Systems
Synthetic Data Generation
Automated Evaluation
Model Benchmarking
Open-source ML Models
Automation & Pipelines
Technical Writing

Job description

24-MAG LLC is seeking a remote Member of Technical Staff in Research Engineering with deep RL expertise. The role spans RL environment design, scalable training pipelines, synthetic data generation, automated evaluation, and benchmarking to accelerate AI research and development.

You will work at the interface of research and production, building RL environments, training workflows, and evaluation systems that improve model capability, reliability, and research velocity in a fully remote,

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote RL Engineer — Scalable ML Systems
Remote RL Engineer — Scalable ML Systems

Bright Vision Technologies • Monroeville

Remote
USD 100,000 - 150,000
Remote | Member of Technical Staff, Research Engineering — $400,000–$800,000/year
Remote | Member of Technical Staff, Research Engineering — $400,000–$800,000/year

24-MAG • United States

Remote
USD 400,000 - 800,000
Remote RL Engineer: Scale & Deploy AI Systems
Remote RL Engineer: Scale & Deploy AI Systems

Bright Vision Technologies • Beaverton (OR)

Remote
USD 80,000 - 100,000
Equal Opportunity Employer
Remote Senior RL Scientist — Agent Design & Training
Remote Senior RL Scientist — Agent Design & Training

Centific Global Solutions, Inc. • United States

On-site
USD 200,000 - 250,000
Senior AI Engineer - Remote Research & RLHF Evaluation
Senior AI Engineer - Remote Research & RLHF Evaluation

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Remote work
Flexible hours
Short onboarding
Remote ML Engineer - RLHF & Model Evaluation Expert
Remote ML Engineer - RLHF & Model Evaluation Expert

prolificacademicltd • United States

Remote
USD 55,000 - 110,000
Competitive hourly rates
Remote-first, flexible schedule
Lightweight onboarding
Remote Research Engineer for AI Code & Model Evaluation
Remote Research Engineer for AI Code & Model Evaluation

24-MAG • United States

Remote
USD 69,000 - 138,000
Remote work
Flexible hours
Contractor engagement
Senior Applied RL Engineer - Remote/Hybrid (US)
Senior Applied RL Engineer - Remote/Hybrid (US)

Centific • Palo Alto (CA), Northern (KY)

Hybrid
USD 150,000 - 300,000
Senior AI/ML Research Engineer (Remote, Flexible Tasks)
Senior AI/ML Research Engineer (Remote, Flexible Tasks)

prolificacademicltd • United States

Remote
USD 66,000 - 110,000
Fully remote engagement
Flexible scheduling
Skills assessment before paid tasks
Senior AI/ML Engineer - LLM Training & Evaluation (Remote)
Senior AI/ML Engineer - LLM Training & Evaluation (Remote)

prolificacademicltd • United States

Remote
USD 69,000 - 110,000
Hourly pay up to $80
Fully remote
Flexible hours
+1