AI Coding Trace Auditor - Rubric-Based Feedback

Mercor

San Francisco (CA)

On-site

USD 120,000 - 180,000

Full time

11 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Mercor in San Francisco seeks an experienced software development evaluator to assess AI-assisted coding traces used to train and evaluate frontier AI models. You will review end-to-end coding sessions produced with AI-assisted developer tools, judging correctness, workflow soundness, and reasoning, and deliver clear rubric-based feedback.

The role emphasizes collaboration with ML researchers and software engineers; you will apply structured criteria, document findings, and help improve tooling

Qualifications

  • 3+ years professional software development.
  • Hands-on experience with AI-assisted coding tools and agentic workflows.
  • Strong code-reading and debugging across full-stack or backend systems.
  • Ability to evaluate multi-step coding trajectories for correctness and best practice.

Responsibilities

  • Evaluate the quality and correctness of AI-assisted software-development traces used to train and evaluate models.
  • Judging correctness, workflow soundness, and reasoning of end-to-end coding sessions.
  • Provide rubric-based written feedback.

Skills

AI-assisted coding tools
Code-reading & debugging
Evaluating multi-step trajectories
Full-stack or backend proficiency

Tools

Cursor
GitHub Copilot
Claude Code
Amazon CodeCatalyst
Kiro

Job description

Mercor in San Francisco seeks an experienced software development evaluator to assess AI-assisted coding traces used to train and evaluate frontier AI models. You will review end-to-end coding sessions produced with AI-assisted developer tools, judging correctness, workflow soundness, and reasoning, and deliver clear rubric-based feedback.

The role emphasizes collaboration with ML researchers and software engineers; you will apply structured criteria, document findings, and help improve tooling

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Code Trace Auditor — Quality Feedback
AI Code Trace Auditor — Quality Feedback

Mercor • United States

Remote
USD 90,000 - 130,000
AI Code Trace Auditor for Developer Tools
AI Code Trace Auditor for Developer Tools

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 170,000
AI Developer Trace Task Auditor
AI Developer Trace Task Auditor

Mercor • San Francisco (CA)

On-site
USD 120,000 - 180,000
AI Developer Trace Task Auditor
AI Developer Trace Task Auditor

Obsidian • San Francisco (CA)

On-site
USD 120,000 - 170,000
AI Developer Trace Task Auditor Mercor · Remote — United States $70-90/hr →
AI Developer Trace Task Auditor Mercor · Remote — United States $70-90/hr →

Dorado • Northern (KY)

Hybrid
USD 90,000 - 130,000
AI-Assisted Code Quality Evaluator
AI-Assisted Code Quality Evaluator

Dorado • Northern (KY)

Hybrid
USD 90,000 - 130,000
Remote AI Code Trace Evaluator (Contract)
Remote AI Code Trace Evaluator (Contract)

OpenTrain AI • Northern (KY)

Hybrid
USD 96,000 - 124,000
C/C++ AI Code Evaluator – Contract
C/C++ AI Code Evaluator – Contract

Turing • San Francisco (CA)

On-site
USD 83,000 - 124,000
AI Benchmarking Engineer — Evaluation & Failure Analysis
AI Benchmarking Engineer — Evaluation & Failure Analysis

Doist • San Francisco (CA)

On-site
USD 150,000 - 210,000
Bi-annual bonus
Equity grant
Relocation bonus
+6
AI Coding Output Reviewer (Multi-Language) — Remote
AI Coding Output Reviewer (Multi-Language) — Remote

AuraOne • United States

On-site
USD 34,440 - 68,880
Remote work