Research Engineer

talentpluto

San Francisco (CA)

On-site

USD 120,000 - 140,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Medical/dental/vision coverage
Meals
401(k)
Commuter benefits
Wellness perk

Job summary

talentpluto is seeking a Research Engineer to enhance the quality assurance (QA) systems supporting training data for reinforcement learning. This position demands close collaboration with stakeholders to guarantee reliability and consistency in datasets.

Key responsibilities include defining quality standards, building auditing workflows, and integrating learnings into processes. Ideal candidates will have strong Python skills, experience with Docker, and a proven ability to operate independently in fast-paced settings. Competitive compensation and benefits are offered.

Qualifications

  • Proficiency with Python and experience working in Linux environments.
  • Experience with Docker and reproducible development/deployment workflows.
  • Experience working with large-scale datasets (validation, transformation, or analysis).
  • Strong problem-solving skills and evidence of rapid learning in technical environments.
  • Ability to operate independently and deliver results in an early-stage, fast-moving setting.

Responsibilities

  • Define and enforce quality standards for training datasets.
  • Build tooling and workflows to audit supplier-generated datasets.
  • Evaluate and implement human-in-the-loop review workflows.
  • Partner with external data suppliers to debug quality issues.
  • Integrate QA learnings into internal tools and supplier portals.

Skills

Proficiency with Python
Experience with Docker
Strong problem-solving skills
Clear written and verbal communication

Job description

Location: San Francisco Bay Area
Work model: On-site (some team members are remote, but this role is currently on-site)
Industry: AI infrastructure / Reinforcement Learning (RL) training data & evaluations
Compensation: Competitive (range not provided) + benefits (medical/dental/vision coverage, meals, 401(k), commuter benefits, wellness perk)

About the Company (our partner)

Our partner is a fast-growing, venture-backed AI infrastructure company building the tooling and workflows that power reinforcement learning (RL) training data and evaluation for frontier AI agents. Their platform is used by advanced AI teams across large enterprises and high-growth startups, and they’re scaling quickly to meet strong customer demand. The team is small, highly technical, and execution-focused, with a culture that values ownership, speed, and craftsmanship.

The Opportunity

Our partner is hiring a Research Engineer to help scale the quality assurance (QA) systems behind training data generated through their infrastructure. This role sits at the intersection of data quality, tooling, and applied ML operations: you’ll build the standards, pipelines, and feedback loops that ensure datasets are reliable, consistent, and ready for training and evaluation.

You’ll work closely with internal stakeholders and external data suppliers to diagnose quality issues, improve workflows, and continuously fold QA learnings back into the platform. If you enjoy building systems that make high-quality data scalable—and want to do it in a high-ownership, fast-paced environment—this role is a strong fit.

Responsibilities
  • Define and enforce quality standards for training datasets used for RL training and evaluation
  • Build tooling and workflows to audit supplier-generated datasets, including sampling strategies, validation pipelines (rule-based and model-assisted), and feedback loops
  • Evaluate and implement human-in-the-loop review workflows where beneficial to improve quality and efficiency
  • Partner with external data suppliers to debug quality issues, provide actionable feedback, and improve their data generation processes
  • Integrate QA learnings into internal tools and supplier portals to reduce anomalies, inconsistencies, and edge cases over time
  • Track QA outcomes and continuously improve processes, metrics, and documentation
Requirements
  • Proficiency with Python and experience working in Linux environments
  • Experience with Docker and reproducible development/deployment workflowsExperience working with large-scale datasets (validation, transformation, or analysis)
  • Strong problem-solving skills and evidence of rapid learning in technical environments
  • Ability to operate independently and deliver results in an early-stage, fast-moving setting
  • Clear written and verbal communication skills (including collaborating across time zones)

Nice to have

  • Experience building data validation pipelines and/or human-in-the-loop review systems
  • Familiarity with common training-data failure modes and techniques to detect subtle inconsistencies
  • Comfort designing QA metrics, experiments, and processes—not just executing predefined checks
  • Familiarity with modern AI tooling and LLM capabilities
Equal Opportunity & Accessibility

Our partner is an Equal Opportunity Employer and is committed to building an inclusive workplace. They consider all qualified applicants without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic. Reasonable accommodations are available throughout the hiring process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer, QC Automation
Research Engineer, QC Automation

Invictus Direct • San Francisco (CA)

On-site
USD 100,000 - 200,000
Relocation assistance
Visa support
Research Engineer: RL Data QA & Tooling
Research Engineer: RL Data QA & Tooling

talentpluto • San Francisco (CA)

On-site
USD 120,000 - 140,000
Medical/dental/vision coverage
Meals
401(k)
+2
Research Engineer, QC Automation
Research Engineer, QC Automation

HUD • San Francisco (CA)

On-site
USD 140,000 - 200,000
Medical, dental, vision coverage (Blue
Lunch & dinner in office
Holiday break & PTO/holidays
+4
Lead Research Engineer, Data Quality
Lead Research Engineer, Data Quality

HUD • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Top-tier medical, dental, vision
Lunch & dinner in office
Holiday break
+4
Software Engineer, Data Quality
Software Engineer, Data Quality

re-zoo-me • San Francisco (CA), Northern (KY)

Hybrid
USD 120,000 - 180,000
Research Engineer
Research Engineer

AfterQuery • San Francisco (CA)

On-site
USD 120,000 - 180,000
Health Insurance: Medical, Vision,/Dan
401(k) with Employer Match
Daily Meals: UberEats stipend
+1
Research Engineer – RL Infrastructure & Agent Environments
Research Engineer – RL Infrastructure & Agent Environments

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 150,000 - 190,000
Software Engineer, Data Quality
Software Engineer, Data Quality

Physical Intelligence • United States

On-site
USD 90,000 - 150,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

On-site
USD 180,000 - 220,000
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

SupportFinity™ • San Francisco (CA)

On-site
USD 100,000 - 180,000