LLM Evaluation Specialist

OpenTrain AI, Inc.

Northern (KY)

Hybrid

USD 45,000 - 65,000

Part time

2 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

OpenTrain AI, Inc. is seeking an LLM Evaluation Specialist to test large language models and assess their performance. You will reason carefully, follow detailed instructions, and review model behavior in a mostly asynchronous, remote setup.

This part-time freelance contractor role pays up to $40 per hour and is open to students or graduates from any discipline. The position requires strong reasoning, attention to detail, and the ability to work independently in a flexible schedule within the

Qualifications

  • Enrolled in an associate, bachelor’s, or master’s program, or completed degree.
  • Strong reasoning, exceptional attention to detail, consistent accuracy.
  • Comfort working independently in an asynchronous environment.
  • Openness to brief training and learning new evaluation workflows.

Responsibilities

  • Test large language models as part of AI training projects.
  • Use general academic knowledge and reasoning to evaluate model performance.
  • Follow detailed instructions while completing evaluation tasks.
  • Review model behavior carefully to support better system performance.

Skills

Strong reasoning
Attention to detail
Independent work
Asynchronous work

Education

Currently enrolled in degree program or completed degree

Job description

The work

As an LLM Evaluation Specialist, you will test large language models and assess how well they perform. You will use careful reasoning, follow detailed instructions, and review model behavior accurately while working independently in a mainly asynchronous setting.

  • Test large language models as part of AI training projects.
  • Use general academic knowledge and reasoning to evaluate model performance.
  • Follow detailed instructions while completing evaluation tasks.
  • Review model behavior carefully and provide work that can support better system performance.
  • Work independently in a remote, primarily asynchronous environment.
What it pays and takes

This is a part-time freelance contractor role for short-term AI training projects. The work is suitable for students and graduates from any academic discipline, and no specific professional or AI training background is required.

  • Pay: Up to $40 per hour.
  • Work type: Remote, part-time freelance contract work.
  • Location: United States.
  • Language: English.
  • Hours: The structured role details list 20+ hours per week; the role description says there is no minimum weekly commitment.
  • Education: Current enrollment in an associate, bachelor's, or master's program, or completion of a degree.
  • Requirements: Strong reasoning, exceptional attention to detail, consistent accuracy, and comfort working independently.
  • Training: Openness to brief training and learning new evaluation workflows.
About AI training work

AI training is the human work behind systems such as language models, including testing responses and rating model performance. OpenTrain hires and contracts contributors for this work, and people with strong judgment are paid to help make AI systems more useful and reliable.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote LLM Evaluation Specialist - Part-Time
Remote LLM Evaluation Specialist - Part-Time

OpenTrain AI, Inc. • Northern (KY)

Hybrid
USD 45,000 - 65,000
Code Quality Engineer for LLM Evaluation
Code Quality Engineer for LLM Evaluation

OpenTrain AI, Inc. • United States

Remote
USD 34,000 - 69,000
Remote AI/ML Engineer for LLM Alignment & Evaluation
Remote AI/ML Engineer for LLM Alignment & Evaluation

prolificacademicltd • United States

Remote
USD 55,000 - 110,000
Hourly pay up to $80
AI QA Trainer - LLM Evaluation - Freelance Project
AI QA Trainer - LLM Evaluation - Freelance Project

Meridial • United States

On-site
USD 8,265 - 89,544
Secure computer and high-speed internet required
LLM AI Training & Evaluation Engineer (Remote)
LLM AI Training & Evaluation Engineer (Remote)

Prolific Academic Ltd • United States

Remote
USD 83,000 - 110,000
Flexible hours
Work from home
Competitive pay up to $80/hr
Legal AI Response Evaluation Lawyer
Legal AI Response Evaluation Lawyer

OpenTrain AI, Inc. • Northern (KY)

Hybrid
USD 124,000 - 193,000
Senior AI Training Engineer - LLM Evaluation & RLHF
Senior AI Training Engineer - LLM Evaluation & RLHF

Prolific Academic Ltd • United States

Remote
USD 83,000 - 138,000
Flexible hours
Work from home
Competitive pay
+1
Remote AI/ML Engineer for LLM Training & Evaluation
Remote AI/ML Engineer for LLM Training & Evaluation

Prolific Academic Ltd • United States

Remote
USD 66,000 - 110,000
Remote AI/ML Engineer — LLM Training & Evaluation
Remote AI/ML Engineer — LLM Training & Evaluation

Rex.zone • United States

Remote
USD 41,328 - 68,880
Senior AI/ML Engineer — LLM Training & Evaluation (Remote)
Senior AI/ML Engineer — LLM Training & Evaluation (Remote)

Prolific Academic Ltd • United States

Remote
USD 66,000 - 110,000