LATAM Software Engineers: Coding Tasks for AI Evaluation

Terac

Argentina

A distancia

ARS 69.247.000 - 98.625.000

A tiempo parcial

Hace 3 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Una candidatura hecha para este puesto de trabajo — un currículum y una carta de presentación adaptados que responden directamente a la oferta.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Hourly pay

Descripción de la vacante

Terac is conducting a paid study on programming task evaluation for AI agents. We are seeking software engineers in Brazil or Argentina to review tasks, verify harness logic, and provide feedback on realism and complexity.

You will share your screen during sessions, discuss architectures, and help improve task structures. Compensation is $40 per hour, with remote participation across Brazil and Argentina.

Formación

  • Professional experience as a software engineer.
  • Located in Brazil or Argentina.
  • Familiarity with building or verifying coding tasks and test environments.
  • Comfortable discussing technical architectures and evaluation frameworks.

Responsabilidades

  • Review realistic programming tasks designed for AI evaluation.
  • Assess the accuracy and structure of various evaluation harnesses.
  • Walk us through your thought process while checking code validity.
  • Provide actionable feedback on how to improve task complexity and realism.

Conocimientos

Software engineer
Code review
Evaluation frameworks
Technical discussion

Descripción del empleo

What We\u2019re Researching

We\u2019re running a paid study on the effectiveness of coding tasks used to evaluate AI agents. Our team is building a comprehensive suite of programming environments designed to test complex software capabilities. We want to ensure these evaluation frameworks are realistic, accurate, and properly calibrated.

How It Works

You will review a series of proposed coding tasks and assess their suitability for testing AI agents. We will ask you to verify the logic of the evaluation harnesses and provide feedback on their realistic application. You will share your screen to walk through the environments and point out potential flaws or improvements. The session involves a mix of code review, technical discussion, and direct feedback on the task structures.

Who This Is For

We are looking for software engineers based in Brazil and Argentina with strong technical backgrounds. You should have direct experience building, reviewing, or testing evaluation harnesses and programming tasks. We welcome backend developers, full-stack engineers, and quality assurance automation specialists.

What You\u2019ll Do
  • Review realistic programming tasks designed for AI evaluation
  • Assess the accuracy and structure of various evaluation harnesses
  • Walk us through your thought process while checking code validity
  • Provide actionable feedback on how to improve task complexity and realism
Who Should Apply
  • Professional experience as a software engineer or developer
  • Located in Brazil or Argentina
  • Familiarity with building or verifying coding tasks and test environments
  • Comfortable discussing technical architectures and evaluation frameworks
Compensation

$40 per hour

About Terac

Terac is building the world\u2019s largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.

Learn more at terac.com or on YouTube at @jointerac.

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Software Engineer, AI Evaluation Harness Reviewer
Software Engineer, AI Evaluation Harness Reviewer

Terac • Argentina

A distancia
ARS 69.247.000 - 98.625.000
Hourly pay
Senior Software Engineer, AI Training - Argentina
Senior Software Engineer, AI Training - Argentina

G2i Inc. • Argentina

A distancia
ARS 209.054.000 - 418.107.000
Software Engineering AI Trainer
Software Engineering AI Trainer

Chromestarchemicals • Comisión de Fomento de Perú

Presencial
ARS 125.533.000 - 251.066.000
Senior AI Coding Evaluator: Engineering Taste (Remote)
Senior AI Coding Evaluator: Engineering Taste (Remote)

G2i Inc. • Argentina

A distancia
ARS 209.054.000 - 418.107.000
AI Agent Evaluation Engineer - Project-Based
AI Agent Evaluation Engineer - Project-Based

Mindrift • Argentina

Presencial
ARS 49.308.150 - 70.226.760
Remote AI Software Engineer Trainer & Task Designer
Remote AI Software Engineer Trainer & Task Designer

Chromestarchemicals • Comisión de Fomento de Perú

Híbrido
ARS 125.533.000 - 251.066.000
Remote AI Software Engineer Trainer — Part-Time (Argentina)
Remote AI Software Engineer Trainer — Part-Time (Argentina)

Anyone AI • Municipio de Esquel

Presencial
ARS 61.759.167 - 144.104.725
Remote Backend Engineer - Contract Task Designer & Evaluator
Remote Backend Engineer - Contract Task Designer & Evaluator

Anyone AI Inc. • Argentina

A distancia
ARS 52.486.000 - 125.965.000
AI Training Contributor - Spanish (Latin America) - Remote
AI Training Contributor - Spanish (Latin America) - Remote

LILT, Inc. • Argentina

A distancia
ARS 51.974.000 - 103.948.000
Flexible schedule
Competitive rates
Global projects
+2
Backend Engineer for AI Code Evaluation - Remote
Backend Engineer for AI Code Evaluation - Remote

Gramian Consulting Group • Argentina

A distancia
ARS 189.540.000 - 273.780.000