Research Engineer - Scalable QC for RL Training Data

HUD

San Francisco (CA)

On-site

USD 140,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision coverage (Blue
Lunch & dinner in office
Holiday break & PTO/holidays
Equinox membership
401k
Commuter benefits
Unlimited AI tool tokens

Job summary

HUD is seeking Research Engineers to automate QC for training data generated through our ML infrastructure, building systems that scale quality for RL training. You’ll implement robust data validation pipelines, sampling strategies, and feedback loops while collaborating with data vendors.

Ideal candidates have strong Python, Docker, and Linux skills, startup experience, and a knack for designing metrics and experiments to assess data quality and task realism.

Qualifications

  • Experience building scalable QA/QC pipelines for datasets.
  • Ability to define quality standards for training data.
  • Experience with benchmarks and RL evals.
  • Startup environment and independent work ability.
  • Strong communication across time zones.

Responsibilities

  • Create QC systems based on true understanding and human judgement, without relying heavily on LLMs.
  • Define and enforce quality standards for training data.
  • Design experiments and metrics to grade agent outputs.
  • Partner with data vendors to debug quality issues and diagnose agent failure modes, provide actionable feedback, and improve their data generation processes
  • Translate QC learnings into systems for auditing supplier-generated datasets, including sampling strategies, validation pipelines (rule-based and model-assisted), and feedback loops
  • Continuously integrate QC learnings into infrastructure tools and data vendor portal to reduce anomalies, inconsistencies, and edge cases

Skills

Python
Docker
Linux
Quality Assurance
RL benchmarks

Education

Bachelor's degree in CS/CE/Math

Tools

Git
Airflow
CI/CD

Job description

HUD is seeking Research Engineers to automate QC for training data generated through our ML infrastructure, building systems that scale quality for RL training. You’ll implement robust data validation pipelines, sampling strategies, and feedback loops while collaborating with data vendors.

Ideal candidates have strong Python, Docker, and Linux skills, startup experience, and a knack for designing metrics and experiments to assess data quality and task realism.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer: RL Data QA & Tooling
Research Engineer: RL Data QA & Tooling

talentpluto • San Francisco (CA)

On-site
USD 120,000 - 140,000
Medical/dental/vision coverage
Meals
401(k)
+2
Research Engineer, QC Automation
Research Engineer, QC Automation

HUD • San Francisco (CA)

On-site
USD 140,000 - 200,000
Medical, dental, vision coverage (Blue
Lunch & dinner in office
Holiday break & PTO/holidays
+4
Research Engineer
Research Engineer

talentpluto • San Francisco (CA)

On-site
USD 120,000 - 140,000
Medical/dental/vision coverage
Meals
401(k)
+2
Research Engineer — RL Training Data & Evals
Research Engineer — RL Training Data & Evals

HUD • San Francisco (CA)

On-site
USD 120,000 - 180,000
Blue Shield medical, dental, vision
Lunch and dinner in office
Holiday break
+4
Lead Research Engineer, Data Quality
Lead Research Engineer, Data Quality

PVH (Tommy Hilfiger/Calvin Klein) • San Francisco (CA)

Hybrid
USD 190,000 - 280,000
Medical, dental, vision (100% covered)
Lunch and dinner in the office
Equinox membership
+3
Research Engineer (General)
Research Engineer (General)

HUD • San Francisco (CA)

On-site
USD 120,000 - 180,000
Blue Shield medical, dental, vision
Lunch and dinner in office
Holiday break
+4
Applied AI Engineer - Deploy, Debug & Build RL Evals
Applied AI Engineer - Deploy, Debug & Build RL Evals

HUD • San Francisco (CA)

On-site
USD 180,000 - 260,000
Medical, dental, vision benefits
Meals in office
Holiday break: Christmas Eve to New Y
+4
Research Engineer - Frontier AI Training & Evals
Research Engineer - Frontier AI Training & Evals

Visa Hunt • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 200,000
Competitive compensation
Medical, dental, vision coverage
Lunch and dinner in office
+5
Research Software Engineer: Scalable RL Training Systems
Research Software Engineer: Scalable RL Training Systems

Jobzhr • New York (NY)

On-site
USD 180,000 - 240,000
Salary and equity
Stock options
Healthcare
+3
Member of Technical Staff, RL Infra
Member of Technical Staff, RL Infra

Inception • San Francisco (CA)

On-site
USD 180,000 - 240,000