Content Evaluator – Bilingual (Vietnamese and English)- Flexible Hours

Innodata Inc.

United States

Hybrid

USD 23,000 - 32,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Innodata Inc. in the United States is seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on Instagram, WhatsApp, and Messenger.

You will evaluate and benchmark AI model responses against complex rubrics using provided business knowledge bases. This full-time, two-month role requires strong written English and Vietnamese, a customer service background, and the ability to follow multi-tier guidelines.

Qualifications

  • Must have customer service experience across email, chat, or phone
  • Excellent written English with tone and grammar accuracy
  • Excellent written Vietnamese with tone and nuance
  • Ability to follow multi-tier evaluation guidelines and rubrics precisely
  • Comfort using web-based labeling interfaces and tools

Responsibilities

  • Review and score AI-generated customer interactions across Foundational, Experiential, and Operational dimensions
  • Benchmark informational and transactional queries against authoritative sources within the task UI
  • Participate in dual-review processes and calibration audits to ensure inter-rater alignment
  • Deliver precise evaluations in a timely manner

Skills

Customer service
English proficiency
Vietnamese proficiency
Analytical precision
Tech adaptability

Job description

Innodata(Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked.Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale.We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.

Role Overview

We are seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on Instagram, WhatsApp, and Messenger. You will evaluate and benchmark AI model responses against complex evaluation rubrics using provided business knowledge bases.

What You’ll Own:
  • Model Evaluation: Review and score AI-generated customer interactions across Foundational, Experiential, and Operational dimensions (e.g., Action Fidelity, Faithfulness, Hallucination, Compliance, Tone, and Handoff).
  • Intent & Fact Verification: Benchmark both informational (R1) and transactional (R2) customer queries against authoritative business sources (FAQs, product catalogs, SOPs) within the task UI.
  • Quality Assurance: Participate in dual-review processes and daily calibration audits to ensure inter-rater agreement and establish ground-truth performance targets.
  • Performance Targets: Deliver precise evaluation
You’ll Thrive in This Role If You Have:
  • Customer Service Background: Prior experience in customer service, call centers, retail, or handling customer communications via email, chat, or phone (highly prioritized).
  • English Proficiency: Exceptional written English skills with a strong command of tone, brand voice, grammar, and nuance.
  • Vietnamese Proficiency: Exceptional written Vietnamese skills with a strong command of tone, brand voice, grammar, and nuance.
  • Analytical Precision: Ability to strictly follow multi-tier evaluation guidelines, complex logic trees, and technical rubrics without deviation.
  • Tech Adaptability: Comfort using dedicated web-based tools and labeling interfaces.
Work Environment & Schedule:

Duration: Full-Time, 2-Month

The expected hourly salary range for this position is up to $20 p/hour, based on experience, skills, and qualifications.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Content Evaluator – Bilingual (French and English)- Flexible Hours
Content Evaluator – Bilingual (French and English)- Flexible Hours

Innodata Inc. • United States

Hybrid
USD 60,000 - 90,000
Bilingual Content Evaluator (Vietnamese & English) for AI QA
Bilingual Content Evaluator (Vietnamese & English) for AI QA

Innodata Inc. • United States

Hybrid
USD 23,000 - 32,000
Content Evaluator – Bilingual (French and English)- Flexible Hours
Content Evaluator – Bilingual (French and English)- Flexible Hours

Innodata Inc. • Northern (KY)

Hybrid
USD 60,000 - 85,000
Content Evaluator – Bilingual (French and English)- Flexible Hours
Content Evaluator – Bilingual (French and English)- Flexible Hours

Dorado • United States

Hybrid
USD 40,000 - 60,000
Bilingual Content Evaluator (EN/VN) — Remote, Flexible Hours
Bilingual Content Evaluator (EN/VN) — Remote, Flexible Hours

Innodata Inc. • Northern (KY)

Hybrid
USD 25,000 - 30,000
Flexible hours
Remote work (US)
Bilingual AI Content Evaluator — Quality & Compliance
Bilingual AI Content Evaluator — Quality & Compliance

Dorado • United States

Hybrid
USD 40,000 - 60,000
Bilingual AI Content Evaluator (English/French)
Bilingual AI Content Evaluator (English/French)

Innodata Inc. • United States

Hybrid
USD 60,000 - 90,000
AI Voice Evaluation Specialist
AI Voice Evaluation Specialist

Innodata Inc. • Maryland

On-site
USD 28,000 - 37,000
AI Voice Evaluation Specialist
AI Voice Evaluation Specialist

Innodata Inc. • South Carolina

On-site
USD 28,000 - 37,000
AI Voice Evaluation Specialist
AI Voice Evaluation Specialist

Innodata Inc. • Massachusetts

On-site
USD 28,000 - 37,000