Research Engineer: RL Data QA & Tooling

talentpluto

San Francisco (CA)

On-site

USD 120,000 - 140,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Medical/dental/vision coverage
Meals
401(k)
Commuter benefits
Wellness perk

Job summary

talentpluto is seeking a Research Engineer to enhance the quality assurance (QA) systems supporting training data for reinforcement learning. This position demands close collaboration with stakeholders to guarantee reliability and consistency in datasets.

Key responsibilities include defining quality standards, building auditing workflows, and integrating learnings into processes. Ideal candidates will have strong Python skills, experience with Docker, and a proven ability to operate independently in fast-paced settings. Competitive compensation and benefits are offered.

Qualifications

  • Proficiency with Python and experience working in Linux environments.
  • Experience with Docker and reproducible development/deployment workflows.
  • Experience working with large-scale datasets (validation, transformation, or analysis).
  • Strong problem-solving skills and evidence of rapid learning in technical environments.
  • Ability to operate independently and deliver results in an early-stage, fast-moving setting.

Responsibilities

  • Define and enforce quality standards for training datasets.
  • Build tooling and workflows to audit supplier-generated datasets.
  • Evaluate and implement human-in-the-loop review workflows.
  • Partner with external data suppliers to debug quality issues.
  • Integrate QA learnings into internal tools and supplier portals.

Skills

Proficiency with Python
Experience with Docker
Strong problem-solving skills
Clear written and verbal communication

Job description

talentpluto is seeking a Research Engineer to enhance the quality assurance (QA) systems supporting training data for reinforcement learning. This position demands close collaboration with stakeholders to guarantee reliability and consistency in datasets.

Key responsibilities include defining quality standards, building auditing workflows, and integrating learnings into processes. Ideal candidates will have strong Python skills, experience with Docker, and a proven ability to operate independently in fast-paced settings. Competitive compensation and benefits are offered.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer
Research Engineer

talentpluto • San Francisco (CA)

On-site
USD 120,000 - 140,000
Medical/dental/vision coverage
Meals
401(k)
+2
Research Engineer - Scalable QC for RL Training Data
Research Engineer - Scalable QC for RL Training Data

HUD • San Francisco (CA)

On-site
USD 140,000 - 200,000
Medical, dental, vision coverage (Blue
Lunch & dinner in office
Holiday break & PTO/holidays
+4
Research Engineer, QC Automation
Research Engineer, QC Automation

Invictus Direct • San Francisco (CA)

On-site
USD 100,000 - 200,000
Relocation assistance
Visa support
AI Training Data QC Automation Engineer
AI Training Data QC Automation Engineer

Invictus Direct • San Francisco (CA)

On-site
USD 100,000 - 200,000
Relocation assistance
Visa support
Full-Stack RL Research Tools Engineer
Full-Stack RL Research Tools Engineer

Cursor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Member of Technical Staff, RL Infra
Member of Technical Staff, RL Infra

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000
Software Engineer, ML Research Tools
Software Engineer, ML Research Tools

Cursor • San Francisco (CA)

On-site
USD 150,000 - 210,000
Head of Quality & MT Staff — RL Verification Lead
Head of Quality & MT Staff — RL Verification Lead

Plato • San Francisco (CA)

On-site
USD 180,000 - 240,000
Reinforcement learning engineer
Reinforcement learning engineer

Dexmate • United States

On-site
USD 100,000 - 130,000
Data Quality Engineer — LLM Post-Training & QA Pipelines
Data Quality Engineer — LLM Post-Training & QA Pipelines

Reflection AI Ltd • New York (NY)

On-site
USD 140,000 - 190,000
Top-tier compensation
Stock options
Health & wellness
+5