AI Session Annotator

Recruitment Room

United States

On-site

USD 50,000 - 62,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Appsierra Group is seeking an AI Session Annotator for a fully remote, independent contractor role. You will review recorded AI sessions, compare responses to visuals, and provide precise written feedback using a detailed rubric.

Strong English, attention to detail, and the ability to follow scoring rubrics are essential; an Android smartphone with camera and mic is required. Work is remote and flexible, with weekly payment.

Qualifications

  • Native or native-level English proficiency, spoken and written.
  • Excellent written English and precise communication.
  • Ability to follow a detailed evaluation rubric.

Responsibilities

  • Review recorded sessions where users interact with an AI assistant through voice and camera.
  • Compare the assistant’s responses with the visual information actually available in the video.
  • Identify fabricated, hallucinated, or inaccurate visual details.
  • Evaluate each session for: visual interpretation, accuracy, relevance, response timing, and use of spatial language.
  • Assess whether the assistant provides appropriate warnings when safety or health could be affected.
  • Write detailed comments supporting each score with specific evidence from the session.
  • Document conversations turn by turn using the intake form and apply the rubric consistently.

Skills

Native English proficiency
Written English
Attention to detail
Analytical judgment
Follow evaluation rubric

Job description

AI Session Annotator

Remote | Independent Contractor | $49,920–$62,400 annualized ($24–$30/hour)

About the Role

Help improve the next generation of voice-and-camera AI assistants by evaluating how accurately they understand and respond to real-world visual information.

As an AI Session Annotator, you’ll review recorded interactions between users and an AI assistant. You’ll compare the assistant’s responses with what was actually visible in the video, evaluate its performance against a detailed rubric, and provide precise written feedback.

Your comments and evaluations will directly help identify issues such as visual hallucinations, inaccurate interpretations, poor timing, and inadequate safety warnings.

This is a judgment-focused role where accuracy, consistency, and quality of feedback matter more than review speed.

What You’ll Do
  • Review recorded sessions where users interact with an AI assistant through voice and camera.

  • Compare the assistant’s responses with the visual information actually available in the video.

  • Identify fabricated, hallucinated, or inaccurate visual details.

  • Evaluate each session for:

    • Visual interpretation

    • Accuracy

    • Relevance

    • Response timing

    • Appropriate use of spatial and directional language

  • Assess whether the assistant provides appropriate and timely warnings when responses could affect a user's health, physical safety, or financial security.

  • Write detailed comments supporting each score with specific evidence from the session.

  • Document conversations turn by turn using the provided intake form.

  • Apply the evaluation rubric consistently without introducing personal scoring criteria.

What You Bring
  • Native or native-level English proficiency, both spoken and written.

  • Excellent written English and the ability to communicate observations clearly and precisely.

  • Strong attention to detail and analytical judgment.

  • Ability to follow a detailed evaluation rubric consistently.

  • Ability to distinguish between what is actually visible in a video and what an AI assistant claims to see.

  • An Android smartphone with a working camera and microphone — mandatory for this project.

Preferred Qualifications

Experience in any of the following is a plus:

  • Data annotation

  • Linguistic quality assurance

  • Content moderation

  • AI model evaluation

  • Linguistics

  • Translation or interpreting

  • Conversational AI evaluation

An ear for regional English usage and the ability to recognize unnatural or non-native phrasing is also valuable.

Content & Project Details
  • Start Date: Immediate

  • High-Volume Period: August 26–31

  • Compensation: $24–$30/hour

  • Annualized Equivalent: $49,920–$62,400

  • Work Arrangement: Fully remote

  • Engagement: Independent contractor

  • Schedule: Flexible, based on project requirements

Annualized compensation is based on 2,080 hours per year for comparison purposes only. Actual earnings depend on the number of hours and projects completed.

Content Considerations

Some sessions involve scenarios where an incorrect AI response could have meaningful real-world consequences. Examples may include identifying medication, determining whether food has spoiled, navigating a street-crossing situation, or interpreting bank card information.

There is no graphic or violent content, but some sessions involve safety-sensitive situations. You’ll be expected to evaluate whether the AI assistant recognized potential risks and provided an appropriate warning when necessary.

Why Join?
  • Help improve the accuracy and reliability of advanced multimodal AI systems.

  • Apply your language and analytical skills to real AI evaluation work.

  • Work remotely with a flexible schedule.

  • Contribute directly to identifying and correcting AI visual reasoning failures.

  • Develop hands-on experience in AI model evaluation and data quality.

Contract & Payment Terms
  • You will be engaged as an independent contractor.

  • Work is fully remote and can be completed on your own schedule.

  • Projects may be extended, shortened, or concluded early depending on project needs and performance.

  • Your work will not require access to confidential or proprietary information belonging to any employer, client, or institution.

  • Payments are made weekly via Stripe or Wise based on services rendered.

  • H-1B and STEM OPT candidates cannot be supported at this time.

Equal Opportunity

All qualified applicants will be considered without regard to legally protected characteristics. Reasonable accommodations are available upon request.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Content Specialist III | AI Evaluation & Prompting
Content Specialist III | AI Evaluation & Prompting

Linda Werner & Associates • United States

Remote
USD 55,000 - 83,000
Health insurance
Health savings account
Life insurance
+2
AI Evaluation Specialist
AI Evaluation Specialist

Weekday AI • United States

Remote
USD 83,000 - 110,000
Ai Session Reviewer - English Expert
Ai Session Reviewer - English Expert

Mercor • San Francisco (CA)

Hybrid
GBP 25,000 - 31,000
AI Agent Evaluation Analyst
AI Agent Evaluation Analyst

Mindrift • Dallas (TX)

On-site
USD 55,104 - 75,768
Flexible remote work
Competitive pay up to $55/hour
Experience in advanced AI projects
AI Agent Evaluation Analyst (Freelance)
AI Agent Evaluation Analyst (Freelance)

Mindrift • Alabama

On-site
USD 90,921 - 129,494
Flexible working hours
Competitive pay up to $80/hour
Experience in advanced AI projects
Video Annotation Specialist | Remote - Contract
Video Annotation Specialist | Remote - Contract

Xperteez Technology Pvt Ltd • United States

Remote
USD 17,000 - 34,000
Video Data Annotator
Video Data Annotator

Remotebridge • United States

Remote
USD 21,000 - 34,000
AI Response Evaluation Analyst
AI Response Evaluation Analyst

iMerit Inc. • United States

On-site
USD 40,000 - 57,000
Fully remote
Spanish Annotation Expert - Fully Remote | Upto $18/hr
Spanish Annotation Expert - Fully Remote | Upto $18/hr

mercor • San Francisco (CA)

Hybrid
MXN 303,000 - 455,000
AI Data Annotation Expert
AI Data Annotation Expert

AuraOne • United States

On-site
USD 34,000 - 62,000