Thai Bilingual Evaluation Specialist (Remote)

AI Trainer Jobs

United States

Remote

USD 34,000 - 55,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

AuraOne is seeking a Thai bilingual expert to evaluate generalist evaluation prompts and responses remotely. You will compare outputs, label edge cases, and produce structured feedback to retrain models. This contractor role supports auditability of labels and rationale for QA.

Applicants should have linguistic or trust & safety review background and be able to work asynchronously. Responsibilities include rating model outputs, tagging content with policy categories, and calibrating against gold

Qualifications

  • Prior evaluation, annotation, or human-rater experience on thai generalist evaluation or adjacent content for Thai Bilingual Expert work.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Experience with inter-rater agreement metrics and calibration cycles.
  • Background in linguistics, content moderation, or trust & safety review.
  • Reliable async availability for at least 10 hours per week.

Responsibilities

  • Evaluate thai generalist evaluation model outputs against a versioned rubric and assign severity tags for Thai Bilingual Expert assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.

Skills

Thai generalist evaluation
Inter-rater calibration
Rubric-based annotation
Severity tagging
Structured feedback writing

Job description

AuraOne is seeking a Thai bilingual expert to evaluate generalist evaluation prompts and responses remotely. You will compare outputs, label edge cases, and produce structured feedback to retrain models. This contractor role supports auditability of labels and rationale for QA.

Applicants should have linguistic or trust & safety review background and be able to work asynchronously. Responsibilities include rating model outputs, tagging content with policy categories, and calibrating against gold

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Thai Evaluation Specialist—Quality & Feedback
Remote Thai Evaluation Specialist—Quality & Feedback

AI Trainer Jobs • United States

Remote
USD 34,000 - 69,000
Remote Vietnamese Bilingual Evaluation Specialist
Remote Vietnamese Bilingual Evaluation Specialist

AuraOne • United States

On-site
USD 28,000 - 55,000
Bilingual Writer - Thai (Thailand)
Bilingual Writer - Thai (Thailand)

AuraOne • United States

On-site
USD 15,000 - 21,000
Remote Bilingual Thai Writer — Evaluation & Feedback
Remote Bilingual Thai Writer — Evaluation & Feedback

AuraOne • United States

On-site
USD 15,000 - 21,000
Thai Language Expert
Thai Language Expert

AI Trainer Jobs • United States

Remote
USD 34,000 - 69,000
Remote Thai Audio Evaluation Specialist – Rubric Review
Remote Thai Audio Evaluation Specialist – Rubric Review

AI Trainer Jobs • United States

Remote
USD 59,000 - 79,000
Remote Bilingual AI Evaluation Specialist
Remote Bilingual AI Evaluation Specialist

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Thai Bilingual Expert
Thai Bilingual Expert

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Remote Tongan Bilingual Evaluation Specialist
Remote Tongan Bilingual Evaluation Specialist

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000
Myanmar Bilingual Evaluation Specialist (Remote)
Myanmar Bilingual Evaluation Specialist (Remote)

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000