Stand out for this role — generate a tailored resume and cover letter in about a minute.
Terac is conducting a remote study to benchmark coding environments used to evaluate AI agents. You will review realistic programming tasks and their evaluation harnesses to ensure they reflect real-world software engineering challenges.
During guided sessions you’ll explain your thought process aloud, participate in screen-sharing exercises, and provide feedback on logic, test cases, and overall environment design. Compensation is $65 per hour.
Terac is conducting a remote study to benchmark coding environments used to evaluate AI agents. You will review realistic programming tasks and their evaluation harnesses to ensure they reflect real-world software engineering challenges.
During guided sessions you’ll explain your thought process aloud, participate in screen-sharing exercises, and provide feedback on logic, test cases, and overall environment design. Compensation is $65 per hour.