Machine Learning Engineer — Model Evaluation & Experimentation

AI Trainer Jobs

United States

Remote

USD 83,000 - 124,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking a Machine Learning Engineer — Model Evaluation & Experimentation contractor to join our remote review team. You will evaluate AI outputs, grade workflow correctness, and ensure policy adherence across reviewer tasks, helping the modeling team improve training data.

In this role, you will apply structured rubrics, document actionable next steps, and communicate findings to stakeholders. Expect 10+ hours per week with flexible scheduling and US-eligible remote work.

Qualifications

  • Direct experience evaluating ML outputs on real teams.
  • Ability to apply multi-page rubrics consistently.
  • Clear written reasoning naming the policy or workflow applied.
  • Reliable async availability for at least 10 hours per week.

Responsibilities

  • Review AI outputs against current workflows, playbooks, and policy.
  • Grade tone, escalation logic, and stakeholder fit on a rubric.
  • Flag operational risk, policy-adherence gaps with severity tags.
  • Capture the right next step so the modeling team can train on it.

Skills

Operational review
Policy adherence
Workflow judgment
Stakeholder communication
Machine learning evaluation

Tools

AI tooling

Job description

Machine Learning Engineer — Model Evaluation & Experimentation is a remote review track for evaluating AI outputs across machine learning engineer model evaluation experimentation specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next step so the modeling team can train on it.

Category: Education & Learning Evaluation · Pay: $60–$90 / hr · Location: Remote — US-eligible · Contractor

Machine Learning Engineer — Model Evaluation & Experimentation is a remote review track for evaluating AI outputs across machine learning engineer model evaluation experimentation specialist operations workflows.

About the role

Machine Learning Engineer — Model Evaluation & Experimentation is a remote review track for evaluating AI outputs across machine learning engineer model evaluation experimentation specialist operations workflows. Reviewers grade workflow correctness, policy adherence, and stakeholder fit; flag operational risk; and document the right next step so the modeling team can train on it.
Machine Learning Engineer Model Evaluation Experimentation specialist operations AI has to fit into an actual day at work. AuraOne uses experienced operators to grade outputs the way a senior peer would — checking workflow, policy, and the unwritten rules that decide whether a task actually gets done.
Review tutoring and assessment output. Judge it against how real learners use it.

Responsibilities
  • Review AI outputs against current machine learning engineer model evaluation experimentation specialist operations workflows, playbooks, and firm policy for Machine Learning Engineer — Model Evaluation & Experimentation assignments.
  • Grade tone, escalation logic, and stakeholder fit on a structured rubric.
  • Flag operational risk, missed escalations, and policy-adherence gaps with severity tags.
  • Capture the right next step so the modeling team can train on it.
Role details

Track Operations & business review Work model Remote · Independent specialist contractor Compensation $60–$90 / hr Eligible from US

What you should bring
  • Direct working experience in machine learning engineer model evaluation experimentation specialist operations on real teams for Machine Learning Engineer — Model Evaluation & Experimentation work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the policy or workflow being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.
Role signals
Example tasks
  • Grade a model's response to a real machine learning engineer model evaluation experimentation specialist operations ticket and rate workflow, tone, and escalation.
  • Flag a missed escalation with the right severity tag and corrected next step.
  • Adjudicate a disputed playbook call between two reviewers using firm guidance.
  • Audit a 25-row batch for rubric consistency and report drift to the program lead.
Useful experience
  • Prior experience training, calibrating, or QA‑ing operations teams.
  • Familiarity with AI-assisted workflow tooling and its failure modes.
  • Bilingual experience for cross‑region operations.
Compensation and schedule

$60–$90 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.

Skills used in matching
  • Operational review
  • Policy adherence
  • Workflow judgment
  • Stakeholder communication
  • Machine Learning Engineer Model Evaluation Experimentation specialist operations
Application boundary

Creating a specialist profile records your experience and preferences. Starting role intake is a separate action that attaches this role to your candidate record.

Specialist intake

The intake preserves your chosen role, the visible terms, and source attribution for reviewer context.
- 01 Confirm profile and eligibility details.
- 02 Attach this role deliberately.
- 03 Receive a human review decision or follow-up.
Placement timing depends on program demand and reviewer confirmation.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Machine Learning Expert
Machine Learning Expert

AI Trainer Jobs • United States

Remote
USD 88,000 - 118,000
Machine Learning Engineer
Machine Learning Engineer

AI Trainer Jobs • United States

Remote
USD 83,000 - 124,000
Machine Learning Engineer Expert
Machine Learning Engineer Expert

AI Trainer Jobs • United States

Remote
USD 105,000 - 143,000
AI & Machine Learning Researcher
AI & Machine Learning Researcher

AI Trainer Jobs • United States

Remote
USD 90,000 - 117,000
Engineering & Software Domain Expert
Engineering & Software Domain Expert

AI Trainer Jobs • United States

Remote
USD 90,000 - 145,000
Member of Technical Staff, AI/ML Engineering
Member of Technical Staff, AI/ML Engineering

AI Trainer Jobs • United States

Remote
USD 179,000 - 207,000
Engineering & Built Environment Experts
Engineering & Built Environment Experts

AI Trainer Jobs • United States

Remote
USD 124,000 - 152,000
Product & Engineering Expert
Product & Engineering Expert

AI Trainer Jobs • United States

Remote
USD 97,000 - 138,000
AI/ML Engineer
AI/ML Engineer

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
Remote work
Evaluation Specialist/Recent Grad
Evaluation Specialist/Recent Grad

AI Trainer Jobs • United States

Remote
USD 99,000 - 135,000