Software Engineers: Paid Code Review for AI Agent Evaluation

AI Trainer Jobs

United States

Remote

USD 83.000 - 96.000

Teilzeit

Vor 8 Tagen
Bewerbungsgenerator

Bekomme eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

AI Trainer Jobs seeks experienced software engineers to participate in a remote, paid study evaluating coding environments and evaluation harnesses used to benchmark AI agents. You will review tasks for realism, verify logic and test cases, and share your reasoning as you walk through complex code structures.

Independent contractor role with weekly payments and flexible scheduling. Join a team that values rigorous code review and practical feedback on building robust evaluation harnesses for AI

Qualifikationen

  • Active professional experience as a software engineer.
  • Hands-on experience building or maintaining test suites and evaluation harnesses.
  • Familiarity with complex real-world software architecture.

Aufgaben

  • Review realistic programming tasks for accuracy and difficulty.
  • Verify the logic and test cases within provided evaluation harnesses.
  • Assess whether coding environments effectively measure software engineering skills.
  • Walk us through your thought process while analyzing complex code structures.

Kenntnisse

Software engineering
Test suites
Evaluation harnesses
Code review

Tools

Test frameworks

Jobbeschreibung

What We're Researching

We're running a paid study on the coding environments and programming tasks used to benchmark artificial intelligence agents. Creating robust evaluation harnesses ensures that AI models are tested against realistic software engineering scenarios. This work directly feeds into improving how autonomous agents handle complex coding objectives.

How It Works

During this remote session, you will review a series of coding tasks and their corresponding evaluation harnesses. You will assess whether the programming challenges accurately reflect real-world software engineering problems. We will ask you to verify the logic, test cases, and overall structure of the environments provided. Your technical feedback will be captured through a guided conversation and screen-sharing exercises.

Who This Is For

We are looking for practicing software engineers with strong backgrounds in building and testing complex systems. Candidates should have direct experience writing test suites, evaluation harnesses, or comprehensive code reviews. We welcome backend engineers, full-stack developers, machine learning engineers, and software architects.

What You'll Do

Review realistic programming tasks for accuracy and difficulty Verify the logic and test cases within provided evaluation harnesses Assess whether coding environments effectively measure software engineering skills Walk us through your thought process while analyzing complex code structures

Who Should Apply

Active professional experience as a software engineer Hands-on experience building or maintaining test suites and evaluation harnesses Comfortable reviewing code and explaining technical concepts aloud Familiarity with complex real-world software architecture

Contract & Payment Terms

You will be engaged as an independent contractor. This is a fully remote opportunity that can be completed on your own schedule. Opportunities can be extended, shortened, or concluded early depending on needs and performance. Your participation will not involve access to confidential or proprietary information from any employer, client, or institution. Payments are processed weekly based on services rendered. We are unable to support H1-B or STEM OPT candidates at this time.

Pay: $65

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Software Engineers: Paid Code Review for AI Agent Evaluation
Software Engineers: Paid Code Review for AI Agent Evaluation

Terac • USA

Remote
USD 138.000 - 413.000
Remote study participation
Compensation for time
South Asian Software Engineers: Coding Tasks for AI Evaluation
South Asian Software Engineers: Coding Tasks for AI Evaluation

AI Trainer Jobs • USA

Remote
USD 70.000 - 76.000
Remote Software Engineer - Evaluation Harness Reviewer
Remote Software Engineer - Evaluation Harness Reviewer

AI Trainer Jobs • USA

Remote
USD 83.000 - 96.000
LATAM Software Engineers: Coding Tasks for AI Evaluation
LATAM Software Engineers: Coding Tasks for AI Evaluation

Remote Jobs • USA

Remote
USD 114.341.000 - 137.760.000
AI Evaluation Engineer — Review Coding Tasks
AI Evaluation Engineer — Review Coding Tasks

Remote Jobs • USA

Remote
USD 114.341.000 - 137.760.000
Software Engineers: Feedback on AI Coding Tools
Software Engineers: Feedback on AI Coding Tools

AI Trainer Jobs • USA

Remote
USD 57 - 69
Agent Engineer
Agent Engineer

AI Trainer Jobs • USA

Remote
USD 138.000 - 689.000
Site Reliability Engineering AI Evaluator
Site Reliability Engineering AI Evaluator

AI Trainer Jobs • USA

Remote
USD 83.000 - 165.000
Remote work
Contractor position
Flexible hours
Senior Software Engineer — AI Coding Evaluator
Senior Software Engineer — AI Coding Evaluator

AI Trainer Jobs • USA

Remote
USD 83.000 - 124.000
Software Engineers: Feedback on AI Coding Tools
Software Engineers: Feedback on AI Coding Tools

Jobgether SRL • USA

Remote
USD 60 - 70
63 USD one-time compensation
Fully remote participation from the US
Contribute to AI research in software