Content Evaluator – Customer Support & LLM Benchmarking

Innodata Inc.

Argentina

Presencial

ARS 29.181.000 - 52.109.000

Jornada completa

Hace 9 días
Generador de candidaturas

Una candidatura completa en un minuto: currículum y carta de presentación adaptados, listos para enviar.

Supera los filtros ATS

Descripción de la vacante

Innodata is seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product across social media platforms. You will evaluate and benchmark AI model responses against complex rubrics using provided knowledge bases.

Hourly commitment is 6 hours per day. This freelance, remote role is for 2 months with potential extension, ideal for candidates with strong English, customer service experience, and meticulous attention to guidelines.

Formación

  • Experience in customer service, call centers, retail, or customer communications.
  • Excellent English proficiency with strong grammar, tone, and comprehension.
  • Strong analytical and attention-to-detail skills.
  • Ability to follow complex guidelines consistently.
  • Comfortable working with web-based evaluation and labeling tools.

Responsabilidades

  • Evaluate AI-generated customer interactions using defined quality parameters.
  • Verify AI responses against FAQs, product catalogs, SOPs, and other approved sources.
  • Participate in dual reviews and daily calibration to maintain evaluation consistency.
  • Follow detailed rubrics and logic trees while meeting productivity targets.
  • Maintain an average handling time of approximately 45 minutes per job.

Conocimientos

Customer service
Analytical skills
Attention to detail
Following guidelines
English proficiency

Herramientas

Web-based evaluation tools

Descripción del empleo

Innodata (NASDAQ: INOD) is a leading data engineering company. With more than 2,000 customers and operations in 13 cities around the world, we are an AI technology solutions provider-of-choice for 4 out of 5 of the world's biggest technology companies, as well as leading companies across financial services, insurance, technology, law, and medicine.

By combining advanced machine learning and artificial intelligence (ML/AI) technologies, a global workforce of subject matter experts, and a high-security infrastructure, we're helping usher in the promise of AI. Innodata offers a powerful combination of both digital data solutions and easy-to-use, high-quality platforms.

Our global workforce includes over 5,000 employees in the United States, Canada, United Kingdom, the Philippines, India, Sri Lanka, Israel and Germany.

About the role:

We are seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on different social media platform. You will evaluate and benchmark AI model responses against complex evaluation rubrics using provided business knowledge bases.

Hourly commitment: 6 hours per day

Key Responsibilities
  • Evaluate AI-generated customer interactions using defined quality parameters.
  • Verify AI responses against FAQs, product catalogs, SOPs, and other approved sources.
  • Participate in dual reviews and daily calibration to maintain evaluation consistency.
  • Follow detailed rubrics and logic trees while meeting productivity targets.
  • Maintain an average handling time of approximately 45 minutes per job.
Required Skills
  • Experience in customer service, call centers, retail, or customer communications.
  • Excellent English proficiency with strong grammar, tone, and comprehension.
  • Strong analytical and attention-to-detail skills.
  • Ability to follow complex guidelines consistently.
  • Comfortable working with web-based evaluation and labeling tools.
  • Duration: 2 months (~60 working days), extendable
  • Schedule: Availability for 6 hours per day
  • Work Type: Freelance

If you're ready to contribute to the future of AI while working remotely on an exciting music-focused project, we'd love to hear from you!

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Remote AI Quality Evaluator & Support for LLM Benchmarking
Remote AI Quality Evaluator & Support for LLM Benchmarking

Innodata Inc. • Argentina

Presencial
ARS 29.181.000 - 52.109.000
Freelance Agent Evaluation Engineer
Freelance Agent Evaluation Engineer

AI Chopping Block • Buenos Aires

Híbrido
ARS 31.166.000 - 62.331.000
Flexible hours
Remote freelance project
Remote Data Scientist: NLP & Production Analytics
Remote Data Scientist: NLP & Production Analytics

IV.AI • Argentina

Presencial
ARS 29.774.000 - 53.593.000
AI Quality Engineer
AI Quality Engineer

Moveo Technologies Corp - drvn • Rosario

Presencial
ARS 1.200.000 - 2.400.000
Contract Operations Systems Analyst
Contract Operations Systems Analyst

Visa Hunt • Argentina

Presencial
ARS 156.004.000 - 176.805.000
Data Analyst
Data Analyst

Firstbase • Buenos Aires

Híbrido
ARS 900.000 - 1.600.000
PTO 15 days
Comprehensive medical coverage
Wellness reimbursement
+2
Data Scientist
Data Scientist

IV.AI • Argentina

A distancia
ARS 29.774.000 - 53.593.000
Data & Machine Learning Engineer
Data & Machine Learning Engineer

IDT • Buenos Aires

Presencial
ARS 104.200.779 - 133.972.431
Senior Technical Support Engineer
Senior Technical Support Engineer

Valid8 Financial, Inc. • Buenos Aires

Híbrido
ARS 104.209.000 - 163.756.000
AI & Support Optimization Analyst
AI & Support Optimization Analyst

Sovos • San Miguel de Tucumán

Presencial
ARS 2.000.000 - 3.500.000
Global Mentoring Programs
AI Tools and Technology from Day 1
Comprehensive Health Insurance
+2