Hebrew Model Evaluation Specialist (Remote)

AI Trainer Jobs

United States

Remote

USD 28,000 - 50,000

Part time

47 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

AuraOne seeks a Hebrew Transcription Expert to remotely evaluate Hebrew outputs using a versioned rubric. You will compare paired responses, label edge cases, and provide structured feedback to retrain the model. This is a contractor role with US-eligibility for remote work.

Responsibilities include evaluating model outputs, tagging issues, and calibrating against gold standards weekly. Strong Hebrew linguistics and attention to detail are required.

Qualifications

  • Prior evaluation, annotation, or human-rater experience on Hebrew evaluation or adjacent content for Hebrew Transcription Expert work.
  • Comfort applying multi-page rubrics consistently across long batches.
  • Clear written reasoning that names the issue and the rubric clause being applied.
  • Strong attention to detail and the ability to flag when a prompt itself is the problem.
  • Reliable async availability for at least 10 hours per week.

Responsibilities

  • Evaluate Hebrew generalist evaluation model outputs against a versioned rubric and assign severity tags for Hebrew Transcription Expert assignments.
  • Compare paired responses and pick the stronger answer with a written rationale.
  • Label hallucinations, instruction-following failures, and unsafe content with structured tags.
  • Capture ambiguous prompts and route them back to the program team for rubric updates.
  • Maintain reviewer-quality scores by calibrating against gold-standard examples each week.

Skills

Model output evaluation
Rubric-based annotation
Severity tagging
Inter-rater calibration
Hebrew generalist evaluation

Job description

AuraOne seeks a Hebrew Transcription Expert to remotely evaluate Hebrew outputs using a versioned rubric. You will compare paired responses, label edge cases, and provide structured feedback to retrain the model. This is a contractor role with US-eligibility for remote work.

Responsibilities include evaluating model outputs, tagging issues, and calibrating against gold standards weekly. Strong Hebrew linguistics and attention to detail are required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Hebrew Language Evaluator — Remote QA & Rubric Review
Hebrew Language Evaluator — Remote QA & Rubric Review

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Remote Hebrew Bilingual Evaluator (Contract)
Remote Hebrew Bilingual Evaluator (Contract)

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000
Remote Hebrew Music & Lyrics Evaluator
Remote Hebrew Music & Lyrics Evaluator

AI Trainer Jobs • United States

Remote
USD 23,000 - 58,000
Hebrew Music Production Evaluation Specialist (Remote)
Hebrew Music Production Evaluation Specialist (Remote)

AI Trainer Jobs • United States

Remote
USD 23,000 - 58,000
Remote work
Hebrew Language Evaluation Specialist (Remote)
Hebrew Language Evaluation Specialist (Remote)

24-Mag Llc • Northern (KY), New York (NY)

Hybrid
USD 41,000 - 90,000
Fully remote
Flexible schedule
Hebrew Transcription Expert
Hebrew Transcription Expert

AI Trainer Jobs • United States

Remote
USD 28,000 - 50,000
Remote Hebrew-English QA & Evaluation Specialist
Remote Hebrew-English QA & Evaluation Specialist

YO AI Labs • Town of Texas (WI)

Remote
USD 34,000 - 58,000
Hebrew Language QA & Linguistic Evaluator — Remote
Hebrew Language QA & Linguistic Evaluator — Remote

YO AI Labs • San Francisco (CA)

Remote
USD 55,000 - 124,000
Remote Hebrew Linguistic Evaluator (Bilingual)
Remote Hebrew Linguistic Evaluator (Bilingual)

YO AI Labs • Austin (TX)

Remote
USD 30,000 - 48,000
Hebrew Bilingual Expert
Hebrew Bilingual Expert

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000