ML Challenge Task Auditor

DigiNo

Northern (KY)

Hybrid

USD 100,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Qualifications

  • 3+ years hands-on applied ML, including experiment design and evaluation.
  • Strong data-quality hygiene: leakage detection and train/test/CV rigor.
  • Proficiency with PyTorch, TensorFlow, scikit-learn, XGBoost.
  • Ability to critique ML claims and reproduce results.

Responsibilities

  • Evaluate quality, correctness, and rigor of applied ML tasks used to train and evaluate models.
  • Assess experiment design, model-selection reasoning, and evaluation methodology.
  • Provide rubric-based written feedback outlining strengths and gaps.

Skills

Applied ML
Experiment design
Model selection
Evaluation methodology
Data quality
ML frameworks

Tools

PyTorch
TensorFlow
scikit-learn
XGBoost

Job description

Evaluate the quality, correctness, and methodological rigor of applied machine-learning tasks used to train and evaluate a frontier AI lab's models. You'll assess experiment design, model-selection reasoning, and evaluation methodology — and provide clear, rubric-based written feedback.

Basic Qualifications
  • 3+ years hands-on applied/experimental ML (experiment design, model selection, hyperparameter tuning, evaluation methodology)
  • Strong grasp of data-quality rigor: leakage detection, metric gaming, and train/test/CV hygiene
  • Proficiency with standard ML frameworks (PyTorch, TensorFlow, scikit-learn, XGBoost)
  • Ability to critique ML claims against evidence and reproduce results
Preferred Qualifications
  • Competition / benchmark experience (e.g., Kaggle)
  • Graduate research or publication record in applied ML
  • Prior task-grading or peer-review experience

Note: this role evaluates applied/experimental ML rigor — it is not an LLM-application-building or MLOps role.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Task Auditor - Applied ML
ML Task Auditor - Applied ML

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 210,000
ML Task Auditor - Applied ML
ML Task Auditor - Applied ML

Mercor • San Francisco (CA)

On-site
USD 130,000 - 190,000
Machine Learning Evaluator - AI Trainer
Machine Learning Evaluator - AI Trainer

Mercor • Philadelphia

On-site
USD 120,000 - 180,000
ML Challenge Task Auditor
ML Challenge Task Auditor

HumanitApp • Northern (KY)

Hybrid
USD 96,000 - 124,000
AI Evaluation Specialist: ML Task Auditor
AI Evaluation Specialist: ML Task Auditor

HumanitApp • Northern (KY)

Hybrid
USD 96,000 - 124,000
Applied ML Task Auditor & Rubric Evaluator
Applied ML Task Auditor & Rubric Evaluator

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 210,000
Applied ML Evaluation Auditor
Applied ML Evaluation Auditor

Obsidian • San Francisco (CA)

Remote
USD 90,000 - 130,000
Applied ML Evaluator & AI Rigor Reviewer
Applied ML Evaluator & AI Rigor Reviewer

Obsidian • Philadelphia

On-site
USD 120,000 - 170,000
ML Rigor Auditor for Experimental Tasks
ML Rigor Auditor for Experimental Tasks

Mercor • San Francisco (CA)

On-site
USD 130,000 - 190,000
Applied ML Evaluator & Methodology Reviewer
Applied ML Evaluator & Methodology Reviewer

Mercor • Philadelphia

On-site
USD 120,000 - 180,000