Applied ML Task Auditor — Experiment Quality Lead

Dorado

Northern (KY)

Hybrid

USD 90,000 - 130,000

Full time

11 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Dorado is seeking a specialist to evaluate the quality, correctness, and methodological rigor of applied ML tasks used to train and evaluate frontier AI lab models. You will assess experiment design, model-selection reasoning, and evaluation methodology, and provide rubric-based written feedback.

The role emphasizes critical appraisal of ML claims, dataset quality, and reproducibility, with no focus on LLM deployment or MLOps.

Qualifications

  • 3+ years hands-on applied ML experience including experiment design, model selection, and evaluation.
  • Strong data-quality rigor including leakage detection, metric gaming, and train/test/CV hygiene.
  • Proficiency with standard ML frameworks: PyTorch, TensorFlow, scikit-learn, XGBoost.
  • Ability to critique ML claims against evidence and reproduce results.

Skills

Hands-on ML experience
Experiment design
Model selection reasoning
Evaluation methodology
Reproducing results
Critiquing ML claims

Tools

PyTorch
TensorFlow
scikit-learn
XGBoost

Job description

Dorado is seeking a specialist to evaluate the quality, correctness, and methodological rigor of applied ML tasks used to train and evaluate frontier AI lab models. You will assess experiment design, model-selection reasoning, and evaluation methodology, and provide rubric-based written feedback.

The role emphasizes critical appraisal of ML claims, dataset quality, and reproducibility, with no focus on LLM deployment or MLOps.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Task Auditor - Applied ML
ML Task Auditor - Applied ML

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 210,000
ML Task Auditor - Applied ML
ML Task Auditor - Applied ML

Mercor • San Francisco (CA)

On-site
USD 130,000 - 190,000
Applied ML Evaluation Auditor
Applied ML Evaluation Auditor

Obsidian • San Francisco (CA)

Remote
USD 90,000 - 130,000
Applied ML Task Auditor & Rubric Evaluator
Applied ML Task Auditor & Rubric Evaluator

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 210,000
ML Challenge Auditor: Rubric-Based Feedback
ML Challenge Auditor: Rubric-Based Feedback

Mercor • United States

Remote
USD 120,000 - 170,000
ML Challenge Task Auditor Mercor · Remote — United States $70-90/hr →
ML Challenge Task Auditor Mercor · Remote — United States $70-90/hr →

Dorado • Northern (KY)

Hybrid
USD 90,000 - 130,000
ML Rigor Auditor for Experimental Tasks
ML Rigor Auditor for Experimental Tasks

Mercor • San Francisco (CA)

On-site
USD 130,000 - 190,000
Applied ML Evaluation Specialist (Remote Contract)
Applied ML Evaluation Specialist (Remote Contract)

OpenTrain AI • Northern (KY)

Hybrid
USD 96,000 - 124,000
Remote ML Task Auditor & AI Quality Engineer
Remote ML Task Auditor & AI Quality Engineer

Meridial • United States

Remote
USD 96,000 - 138,000
SWE Benchmark Auditor for AI Model Evaluation
SWE Benchmark Auditor for AI Model Evaluation

Dorado • Northern (KY)

Hybrid
USD 100,000 - 150,000