Incident management / reliability / SRE Evaluator

AI Trainer Jobs

United States

Remote

USD 110,000 - 165,000

Part time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking a remote Contractor for Incident management / reliability / SRE evaluation. You will reproduce failures, write minimal tests, and explain fixes so the modeling team can target gaps.

Strong coding experience and ability to read unfamiliar code are essential. Responsibilities include grading cloud and platform engineering solutions for correctness, writing failing tests, and ranking paired solutions with a clear rationale.

Qualifications

  • Strong day-job engineering experience with ability to read, run, and debug unfamiliar code.
  • Ability to write concise unit tests capturing a single failure mode.
  • Familiarity with a testing framework (pytest, Jest, JUnit, Go test, RSpec or equivalent).
  • Clear written reasoning to persuade a senior engineer.

Responsibilities

  • Run and reproduce candidate code outputs in sandboxed environments.
  • Grade cloud & platform engineering solutions for correctness, style, and edge-case handling.
  • Write minimal failing tests that demonstrate the bug a model output missed.
  • Compare paired solutions and rank them with a written rationale tied to the rubric.
  • Triage a security issue and document the patch the model should have produced.

Skills

Code review
Debugging
Unit testing
Software engineering judgment
Cloud & platform engineering

Tools

pytest
Jest
JUnit
Go test
RSpec

Job description

Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.

Category: Coding, SWE & Agent Evaluation · Pay: $80–$120 / hr · Location: Remote — US-eligible · Contractor

Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards.

About the role

Incident management / reliability / SRE Evaluator is a remote engineering review track for evaluating production code, debugging traces, and developer-facing AI outputs against real-world correctness standards. Reviewers reproduce failures, write the unit test the model should have written, and explain the fix so the modeling team can target the gap.
Engineering model quality lives or dies on whether the generated code actually compiles, passes tests, and handles edge cases. AuraOne pairs experienced engineers with the modeling team to grade outputs the way a code reviewer would.
Judge generated code and software engineering agents. Read their debugging traces.

Responsibilities
  • Run and reproduce candidate code outputs in a sandboxed environment for Incident management / reliability / SRE Evaluator assignments.
  • Grade cloud & platform engineering solutions for correctness, style, and edge-case handling.
  • Write minimal failing tests that demonstrate the bug a model output missed.
  • Compare paired solutions and rank them with a written rationale tied to the rubric.
Role details

Track Code review & evaluation Work model Remote · Independent specialist contractor Compensation $80–$120 / hr Eligible from US

What you should bring
  • Strong day-job engineering experience — you can read, run, and debug unfamiliar code for Incident management / reliability / SRE Evaluator work.
  • Comfort writing concise unit tests that capture a single failure mode.
  • Familiarity with a testing framework. pytest, Jest, or JUnit. Go test, RSpec, or whatever you use.
  • Clear written reasoning — your review note has to convince another senior engineer.
  • Reliable async availability for at least 10 hours per week.
  • Prior code-review or technical-interview-grading experience is a plus.
Example tasks
  • Reproduce a generated engineering solution to a coding task, run the test suite, and grade it.
  • Write the smallest failing test that demonstrates a model's edge-case bug.
  • Compare two paired solutions and rank them with a written rationale tied to the rubric.
  • Triage a security issue surfaced by a model and document the patch the model should have produced.
Useful experience
  • Open-source contributions or a public portfolio that demonstrates production-quality code.
  • Experience with the target language's standard tooling, linters, and idiomatic style guides.
  • Familiarity with security-review checklists (OWASP, CWE) and AppSec patterns.
Compensation and schedule

$80–$120 / hr
Expected arrangement: contractor , with program-defined task volume and review pacing. Placement depends on current program demand and reviewer confirmation.

Skills used in matching
  • Code review
  • Debugging
  • Unit testing
  • Software engineering judgment
  • Cloud & platform engineering
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineering Evaluation Specialist
Software Engineering Evaluation Specialist

AI Trainer Jobs • United States

Remote
USD 158,000 - 200,000
Software Engineer, Data Platforms
Software Engineer, Data Platforms

AI Trainer Jobs • United States

Remote
USD 165,000 - 193,000
Reliability Engineer
Reliability Engineer

AI Trainer Jobs • United States

Remote
USD 110,000 - 124,000
Systems Engineer (Coding Agent Experience)
Systems Engineer (Coding Agent Experience)

AI Trainer Jobs • United States

Remote
USD 99,000 - 135,000
Remote work
Senior System Engineer
Senior System Engineer

AI Trainer Jobs • United States

Remote
USD 41,000 - 83,000
Remote work
US-eligible contractor
Support Engineer Expert
Support Engineer Expert

AI Trainer Jobs • United States

Remote
USD 153,000 - 205,000
Strategic Project Lead, Coding/SWE
Strategic Project Lead, Coding/SWE

AI Trainer Jobs • United States

Remote
USD 148,000 - 210,000
SecOps Engineer
SecOps Engineer

AI Trainer Jobs • United States

Remote
USD 41,000 - 138,000
DevOps / SRE / Cloud Engineer (Coding Agent Experience)
DevOps / SRE / Cloud Engineer (Coding Agent Experience)

AI Trainer Jobs • United States

Remote
USD 103,000 - 131,000
Remote work
Contractor role
US eligible
Senior Full-Stack Engineer, Multi-Tenant SaaS Platforms
Senior Full-Stack Engineer, Multi-Tenant SaaS Platforms

AI Trainer Jobs • United States

Remote
USD 103,000 - 131,000