Research Manager: Scalable AI Training Data & Evals

HUD

San Francisco (CA)

Hybrid

USD 180,000 - 240,000

Full time

34 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Medical/dental/vision coverage
Lunch and dinner when in office
Holiday break
401k
Commuter benefits
Equinox membership
Unlimited AI tokens

Job summary

HUD is building infrastructure to create RL training data and evals for frontier AI agents and a marketplace to sell them to labs. We seek a Research Manager to lead research engineers, shape data quality methods, and scale insights into repeatable workflows.

You will oversee problem definition, experiments, and conclusions, ensuring impact on agent training while staying close to technical work and cross-team collaboration.

Qualifications

  • Experience leading technical research projects from question to working result.
  • Experience mentoring researchers or research engineers while staying hands-on.
  • Strong understanding of ML and RL and how data shapes model behavior.
  • Experience with agent training data, evals, benchmarks, and validation infra.
  • Sound experimental judgment to distinguish useful signals from misleading metrics.
  • Excellent written communication to explain methods to researchers and partners.

Responsibilities

  • Set research direction for data quality and how tasks, trajectories, rewards, and evals affect training.
  • Lead researchers from problem definition through experiments and conclusions; coach for technical judgment.
  • Design and review experiments linking model behavior to data, env, and reward design.
  • Develop scalable methods to validate and improve training data, including audits and feedback loops.
  • Partner with researchers, domain experts, and vendors to translate insights into workflows and standards.
  • Communicate findings and tradeoffs clearly to help prioritize work across research areas.

Skills

Lead research projects
Mentor researchers
ML / RL expertise
Agent training data & evals
Experimental judgment
Strong written communication

Job description

HUD is building infrastructure to create RL training data and evals for frontier AI agents and a marketplace to sell them to labs. We seek a Research Manager to lead research engineers, shape data quality methods, and scale insights into repeatable workflows.

You will oversee problem definition, experiments, and conclusions, ensuring impact on agent training while staying close to technical work and cross-team collaboration.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Engineer — RL Training Data & Evals
Research Engineer — RL Training Data & Evals

HUD • San Francisco (CA)

On-site
USD 120,000 - 180,000
Blue Shield medical, dental, vision
Lunch and dinner in office
Holiday break
+4
Applied AI Engineer - Deploy, Debug & Build RL Evals
Applied AI Engineer - Deploy, Debug & Build RL Evals

HUD • San Francisco (CA)

On-site
USD 180,000 - 260,000
Medical, dental, vision benefits
Meals in office
Holiday break: Christmas Eve to New Y
+4
Research Engineer - Scalable QC for RL Training Data
Research Engineer - Scalable QC for RL Training Data

HUD • San Francisco (CA)

On-site
USD 140,000 - 200,000
Medical, dental, vision coverage (Blue
Lunch & dinner in office
Holiday break & PTO/holidays
+4
Senior Data & RL Engineer — Coding Agents Lead
Senior Data & RL Engineer — Coding Agents Lead

turing • Palo Alto (CA)

On-site
USD 250,000 - 350,000
Equity
Office in SF/Palo Alto/Seattle
Marketplace Product Manager - AI RL Data (Remote)
Marketplace Product Manager - AI RL Data (Remote)

HUD • San Francisco (CA)

Hybrid
USD 120,000 - 170,000
Competitive compensation
Top-notch medical, dental, vision
Lunch and dinner when in office
+5
Research Engineer (General)
Research Engineer (General)

HUD • San Francisco (CA)

On-site
USD 120,000 - 180,000
Blue Shield medical, dental, vision
Lunch and dinner in office
Holiday break
+4
Senior Data & RL Environments Architect for AI Labs
Senior Data & RL Environments Architect for AI Labs

turing • San Francisco (CA)

On-site
USD 180,000 - 240,000
Lead Research Engineer, Data Quality
Lead Research Engineer, Data Quality

PVH (Tommy Hilfiger/Calvin Klein) • San Francisco (CA)

Hybrid
USD 190,000 - 280,000
Medical, dental, vision (100% covered)
Lunch and dinner in the office
Equinox membership
+3
Forward Deployed Research Engineer
Forward Deployed Research Engineer

HUD • San Francisco (CA)

On-site
USD 180,000 - 260,000
Medical, dental, vision benefits
Meals in office
Holiday break: Christmas Eve to New Y
+4
Research Engineer: Scalable RL & Enterprise AI
Research Engineer: Scalable RL & Enterprise AI

Tessera Labs • New York (NY)

On-site
USD 200,000 - 300,000