SWE Agent Workflow Reviewer

AI Trainer Jobs

United States

Remote

USD 152,000 - 207,000

Full time

2 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking a SWE Agent Workflow Reviewer for a remote independent contractor role. You will reproduce failures, write the unit tests the model should have written, and explain fixes so the modeling team can target the gap.

Engineering model quality hinges on code compiling, tests passing, and edge cases handled in sandboxed environments. AuraOne pairs experienced engineers with the modeling team to grade outputs the way a code reviewer would, judging generated code and software

Qualifications

  • Strong engineering experience to read, run, and debug unfamiliar code.
  • Ability to write concise unit tests capturing a single failure mode.
  • Experience with a testing framework such as pytest, Jest, JUnit, Go test, or RSpec.
  • Clear written reasoning to convince a senior engineer.
  • Reliable async availability for at least 10 hours per week.
  • Prior code-review or technical-interview-grading experience is a plus.

Responsibilities

  • Run and reproduce candidate code outputs in a sandboxed environment.
  • Grade SWE agent workflow solutions for correctness, style, and edge-case handling.
  • Write the smallest failing tests that demonstrate a bug.
  • Compare paired solutions and rank them with a written rubric-based rationale.
  • Triage security issues and document the patch the model should have produced.

Skills

Code review
Debugging
Unit testing
Software testing
Async availability
Open-source contributions
Clear written reasoning

Tools

pytest
Jest
JUnit
Go test
RSpec

Job description

SWE Agent Workflow Reviewer is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.

Category: Coding, SWE & Agent Evaluation · Pay: $130 / hr · Location: Remote — US-eligible · Contractor

SWE Agent Workflow Reviewer is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards.

About the role

SWE Agent Workflow Reviewer is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.
Engineering model quality lives or dies on whether the generated code actually compiles, passes tests, and handles edge cases. AuraOne pairs experienced engineers with the modeling team to grade outputs the way a code reviewer would.
Judge generated code and software engineering agents. Read their debugging traces.

Responsibilities
  • Run and reproduce candidate code outputs in a sandboxed environment for SWE Agent Workflow Reviewer assignments.
  • Grade swe agent workflow engineering review solutions for correctness, style, and edge-case handling.
  • Write minimal failing tests that demonstrate the bug a model output missed.
  • Compare paired solutions and rank them with a written rationale tied to the rubric.
Role details

Track Code review & evaluation Work model Remote · Independent specialist contractor Compensation Hourly rate confirmed after the interview process. Eligible from US

What you should bring
  • Strong day-job engineering experience — you can read, run, and debug unfamiliar code for SWE Agent Workflow Reviewer work.
  • Comfort writing concise unit tests that capture a single failure mode.
  • Familiarity with a testing framework. pytest, Jest, or JUnit. Go test, RSpec, or whatever you use.
  • Clear written reasoning — your review note has to convince another senior engineer.
  • Reliable async availability for at least 10 hours per week.
  • Prior code-review or technical-interview-grading experience is a plus.
Example tasks
  • Reproduce a generated engineering solution to a coding task, run the test suite, and grade it.
  • Write the smallest failing test that demonstrates a model's edge-case bug.
  • Compare two paired solutions and rank them with a written rationale tied to the rubric.
  • Triage a security issue surfaced by a model and document the patch the model should have produced.
Useful experience
  • Open-source contributions or a public portfolio that demonstrates production-quality code.
  • Experience with the target language's standard tooling, linters, and idiomatic style guides.
  • Familiarity with security-review checklists (OWASP, CWE) and AppSec patterns.
Compensation and schedule

Hourly rate confirmed after the interview process.
Expected arrangement: contractor, with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.

Skills used in matching
  • Code review
  • Debugging
  • Unit testing
  • Software engineering judgment
  • SWE Agent Workflow engineering review
  • Software testing
  • Agent evaluation
  • Agent
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Workflow Annotator — Software Engineering
Workflow Annotator — Software Engineering

AI Trainer Jobs • United States

Remote
USD 148,000 - 210,000
Strategic Project Lead, Coding/SWE
Strategic Project Lead, Coding/SWE

AI Trainer Jobs • United States

Remote
USD 148,000 - 210,000
Systems Engineer (Coding Agent Experience)
Systems Engineer (Coding Agent Experience)

AI Trainer Jobs • United States

Remote
USD 99,000 - 135,000
Remote work
Software Engineering Evaluation Specialist
Software Engineering Evaluation Specialist

AI Trainer Jobs • United States

Remote
USD 158,000 - 200,000
Agent Engineer
Agent Engineer

AI Trainer Jobs • United States

Remote
USD 138,000 - 689,000
Incident management / reliability / SRE Evaluator
Incident management / reliability / SRE Evaluator

AI Trainer Jobs • United States

Remote
USD 110,000 - 165,000
Software Engineer, Data Platforms
Software Engineer, Data Platforms

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
Software Engineer, New Grad
Software Engineer, New Grad

AI Trainer Jobs • United States

Remote
USD 148,000 - 210,000
Senior Software Engineer
Senior Software Engineer

AI Trainer Jobs • United States

Remote
USD 148,000 - 210,000
Software Engineer, C — Codebase Q&A
Software Engineer, C — Codebase Q&A

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
Remote work
Contractor role