AI/ML Evaluator - Human-in-the-Loop (Freelance)

Welocalize

Town of Texas (WI)

On-site

USD 192,864,000 - 220,416,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Welocalize is seeking analytical and technically skilled AI/ML Evaluators to review and evaluate complex AI system behaviour using expert human judgment. You will analyse AI system outputs, telemetry and other technical signals to identify meaningful patterns and distinguish legitimate user activity from automated or bot activity.

You'll provide high-quality human evaluations to train, evaluate and improve AI systems, applying your knowledge of AI/ML concepts and detailed project guidelines with

Qualifications

  • Educational background or equivalent experience in AI, ML, CS, data science, statistics, engineering, or related field.
  • Experience with AI/ML systems, data analysis, or technical system analysis.
  • Strong understanding of core AI/ML concepts.
  • Excellent written English and documentation skills.
  • Familiarity with bot detection, anomaly detection, or fraud detection is preferred.
  • Experience with data annotation or model evaluation is preferred.

Responsibilities

  • Review and evaluate AI system behaviour, outputs, and technical data according to project guidelines.
  • Analyse system telemetry, event data, signals, and other technical information to identify meaningful patterns.
  • Evaluate patterns and signals that distinguish legitimate user activity from automated or bot activity.
  • Apply AI/ML knowledge and expert human judgment when reviewing complex or ambiguous cases.
  • Identify unusual patterns, inconsistencies, or behaviours that may require further review.
  • Evaluate complex cases using available evidence, context, and defined project requirements.
  • Classify, label, or annotate assigned data accurately and consistently.
  • Provide high-quality human evaluations to support AI system training, evaluation, and improvement.
  • Document decisions and supporting reasoning clearly when required.
  • Identify unclear or unusual cases and flag them according to defined project processes.
  • Apply detailed project guidelines consistently across assigned tasks.
  • Maintain high levels of quality, accuracy, consistency, and attention to detail.

Skills

Analytical thinking
AI/ML concepts
Data analysis
Pattern recognition
Documentation skills

Education

Bachelor's degree or equivalent in AI/ML/CS/Data Science/Statistics/Engineering

Job description

Welocalize is seeking analytical and technically skilled AI/ML Evaluators to review and evaluate complex AI system behaviour using expert human judgment. You will analyse AI system outputs, telemetry and other technical signals to identify meaningful patterns and distinguish legitimate user activity from automated or bot activity.

You'll provide high-quality human evaluations to train, evaluate and improve AI systems, applying your knowledge of AI/ML concepts and detailed project guidelines with

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI/ML Evaluator - English
AI/ML Evaluator - English

Welocalize • Town of Texas (WI)

On-site
USD 192,864,000 - 220,416,000
AI Evaluation Specialist: ML Task Auditor
AI Evaluation Specialist: ML Task Auditor

HumanitApp • Northern (KY)

Hybrid
USD 96,000 - 124,000
AI Evaluation Specialist
AI Evaluation Specialist

micro1 • United States

On-site
AUD 70,000 - 110,000
Machine Learning (ML) AI Task Auditor - Freelance AI Trainer Project
Machine Learning (ML) AI Task Auditor - Freelance AI Trainer Project

Triwill Group • Northern (KY)

Hybrid
USD 83,000 - 165,000
Remote work
Remote AI Design Evaluator & Creative Specialist
Remote AI Design Evaluator & Creative Specialist

YO AI Labs • San Francisco (CA)

Remote
USD 55,000 - 96,000
Remote Freelance ML Task Auditor & AI Evaluation Specialist
Remote Freelance ML Task Auditor & AI Evaluation Specialist

Triwill Group • Northern (KY)

Hybrid
USD 83,000 - 165,000
Remote work
AI QA Trainer - LLM Evaluation - Freelance Project
AI QA Trainer - LLM Evaluation - Freelance Project

Meridial • United States

Remote
Secure computer and high-speed internet required
Remote: AI Design Evaluator & Creative Specialist
Remote: AI Design Evaluator & Creative Specialist

YO AI Labs • Town of Texas (WI)

Remote
USD 34,000 - 62,000
AI Evaluation Specialist - Remote Contract
AI Evaluation Specialist - Remote Contract

micro1 • United States

Remote
AUD 70,000 - 110,000
Remote Creative & Design Specialist for AI Evaluation
Remote Creative & Design Specialist for AI Evaluation

YO AI Labs • New York (NY)

Remote
USD 55,000 - 124,000