ML Engineer (Coding Agent Experience)

AI Trainer Jobs

United States

Remote

USD 96,000 - 138,000

Part time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking a remote ML Engineer (Coding Agent Experience) to review production code, debugging traces, and judge AI outputs against real-world correctness standards. You will reproduce failures, write minimal failing tests, and explain fixes so the modeling team can target gaps.

This contractor role offers flexible scheduling and US-eligible remote work.

Qualifications

  • Strong day-job engineering experience — read, run, and debug unfamiliar code for ML Engineer (Coding Agent Experience) work.
  • Comfort writing concise unit tests that capture a single failure mode.
  • Familiarity with a testing framework: pytest, Jest, JUnit; Go test, RSpec, or equivalents.
  • Clear written reasoning — your review note must convince a senior engineer.
  • Reliable async availability for at least 10 hours per week.
  • Prior code-review or technical-interview-grading experience is a plus.

Responsibilities

  • Run and reproduce candidate code outputs in a sandboxed environment for ML Engineer (Coding Agent Experience) assignments.
  • Grade ml engineering solutions for correctness, style, and edge-case handling.
  • Write minimal failing tests that demonstrate the bug a model output missed.
  • Compare paired solutions and rank them with a written rationale tied to the rubric.
  • Tag failure modes (compile error, runtime crash, off-by-one, security issue) with severity scores.

Skills

Code reading & debugging
Unit testing
Async availability
Clear written reasoning
ML engineering review experience

Tools

pytest
Jest
JUnit
Go test
RSpec

Job description

ML Engineer (Coding Agent Experience) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.

Category: Coding, SWE & Agent Evaluation · Pay: $85 / hr · Location: Remote — US-eligible · Contractor

ML Engineer (Coding Agent Experience) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards.

About the role

ML Engineer (Coding Agent Experience) is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.
Engineering model quality lives or dies on whether the generated code actually compiles, passes tests, and handles edge cases. AuraOne pairs experienced engineers with the modeling team to grade outputs the way a code reviewer would.
Judge generated code and software engineering agents. Read their debugging traces.

Responsibilities
  • Run and reproduce candidate code outputs in a sandboxed environment for ML Engineer (Coding Agent Experience) assignments.
  • Grade ml engineering solutions for correctness, style, and edge-case handling.
  • Write minimal failing tests that demonstrate the bug a model output missed.
  • Compare paired solutions and rank them with a written rationale tied to the rubric.
  • Tag failure modes (compile error, runtime crash, off‑by‑one, security issue) with severity scores.
Role details

Track Code review & evaluation Work model Remote · Independent specialist contractor Compensation $85 / hr Eligible from US

What you should bring
  • Strong day-job engineering experience — you can read, run, and debug unfamiliar code for ML Engineer (Coding Agent Experience) work.
  • Comfort writing concise unit tests that capture a single failure mode.
  • Familiarity with a testing framework. pytest, Jest, or JUnit. Go test, RSpec, or whatever you use.
  • Clear written reasoning — your review note has to convince another senior engineer.
  • Reliable async availability for at least 10 hours per week.
  • Prior code‑review or technical‑interview‑grading experience is a plus.
Example tasks
  • Reproduce a generated engineering solution to a coding task, run the test suite, and grade it.
  • Write the smallest failing test that demonstrates a model's edge-case bug.
  • Compare two paired solutions and rank them with a written rationale tied to the rubric.
  • Triage a security issue surfaced by a model and document the patch the model should have produced.
Useful experience
  • Open‑source contributions or a public portfolio that demonstrates production‑quality code.
  • Experience with the target language's standard tooling, linters, and idiomatic style guides.
  • Familiarity with security‑review checklists (OWASP, CWE) and AppSec patterns.
Compensation and schedule

$85 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.

Skills used in matching
  • Code review
  • Debugging
  • Unit testing
  • Software engineering judgment
  • ML engineering
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

ML Engineer
ML Engineer

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
AI/ML Engineer
AI/ML Engineer

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
Remote work
Systems Engineer (Coding Agent Experience)
Systems Engineer (Coding Agent Experience)

AI Trainer Jobs • United States

Remote
USD 99,000 - 135,000
Remote work
Member of Technical Staff, AI/ML Engineering
Member of Technical Staff, AI/ML Engineering

AI Trainer Jobs • United States

Remote
USD 179,000 - 207,000
Mechanical Engineer
Mechanical Engineer

AI Trainer Jobs • United States

Remote
USD 124,000 - 220,000
Remote work
Agent Engineer
Agent Engineer

AI Trainer Jobs • United States

Remote
USD 138,000 - 689,000
Support Engineer Expert
Support Engineer Expert

AI Trainer Jobs • United States

Remote
USD 153,000 - 205,000
Data Engineer (Coding Agent Experience)
Data Engineer (Coding Agent Experience)

AI Trainer Jobs • United States

Remote
USD 96,000 - 124,000
Remote work
Mechanical Engineer Expert
Mechanical Engineer Expert

AI Trainer Jobs • United States

Remote
USD 117,000 - 138,000
Remote work
Incident management / reliability / SRE Evaluator
Incident management / reliability / SRE Evaluator

AI Trainer Jobs • United States

Remote
USD 110,000 - 165,000