Expert Prompt Curators for Advanced AI Evaluation Dataset

CloudDevs

United States

On-site

USD 69,000 - 124,000

Part time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Competitive hourly compensation
Set your own hours
Contribution to AI research

Job summary

A leading AI research firm is seeking Expert Prompt Curators to design challenging prompts for evaluating advanced AI models. The role requires advanced knowledge in diverse fields and offers flexible hours, remote work, and a competitive hourly wage. Ideal candidates will have experience in research and test question design. This is a short-term project with potential for extension, contributing to high-impact AI safety research.

Qualifications

  • Advanced academic or professional expertise in a specialized subject.
  • Strong ability to design precise, high-difficulty questions.
  • Experience in academic research, benchmarking, or test question design preferred.
  • Attention to detail and concise reasoning explanations.
  • Familiarity with AI models and their limitations is a plus.

Responsibilities

  • Create original, expert-level prompts that require tool use.
  • Test prompts against advanced AI models and document results.
  • Provide reasoning steps and solutions for each prompt.
  • Collaborate with reviewers for expert validation and refinement.

Skills

Expertise in specialized subjects
Ability to design high-difficulty questions
Attention to detail
Familiarity with AI models
Attention to detail

Education

Advanced academic or professional expertise

Job description

Expert Prompt Curators for Advanced AI Evaluation Dataset

This description is a summary of our understanding of the job description. Click on ‘Apply’ button to find out more.

Role Description

Mercor is collaborating with a leading AI research lab to develop a next-generation evaluation dataset for frontier AI models. We are seeking experts with advanced domain knowledge across diverse fields to design extremely challenging prompts that cannot be solved by existing AI systems without internet search or browsing capabilities. The goal is to create a benchmark dataset that pushes the limits of current AI reasoning and retrieval. This is a short-term research engagement with significant impact on AI evaluation.

Key Responsibilities
  • Create original, expert-level prompts that require tool use (e.g., search, browse, or code execution).
  • Ensure prompts are objective, self-contained, and yield clear, unambiguous answers.
  • Test prompts against advanced AI models and document failures/successes.
  • Provide reasoning steps and solutions for each prompt.
  • Classify prompts into subject domains for dataset organization.
  • Collaborate with reviewers for expert validation and prompt refinement.
Qualifications
  • Advanced academic or professional expertise in a specialized subject (STEM, law, finance, history, cultural studies, etc.).
  • Strong ability to design precise, high-difficulty questions requiring deep knowledge and external references.
  • Experience in academic research, benchmarking, or test question design preferred.
  • Attention to detail and ability to provide concise reasoning explanations.
  • Familiarity with AI models and their limitations is a plus.
Requirements
  • Remote and asynchronous — set your own hours.
  • Expected commitment: ~10–20 hours/week.
  • Project duration: ~2 months, with possible extensions based on dataset needs.
  • Opportunity to contribute to high-impact AI safety and evaluation research.
Compensation & Contract Terms
  • Competitive hourly compensation based on expertise.
  • Independent contractor engagement.
  • Payments for services rendered processed weekly via Stripe Connect.
Application Process
  • Submit your resume or CV highlighting your subject matter expertise.
  • Complete a brief questionnaire about your background and areas of specialization.
  • Selected applicants may be asked to draft a short test prompt.
  • You’ll receive follow-up within a few days regarding next steps.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Expert Prompt Curator for AI Evaluation Dataset
Remote Expert Prompt Curator for AI Evaluation Dataset

CloudDevs • United States

Remote
USD 69,000 - 124,000
Competitive hourly compensation
Set your own hours
Contribution to AI research
Prompt Engineer
Prompt Engineer

Soothsayer Analytics • Livonia (MI)

On-site
USD 60,000 - 80,000
Prompt Engineer
Prompt Engineer

Odiin.AI • New York (NY)

On-site
USD 120,000 - 160,000
Prompt Engineer (AI Strategy & Analytics)
Prompt Engineer (AI Strategy & Analytics)

Mbi Llc • Harrisburg

On-site
USD 85,000 - 110,000
Business Operations Expert - Evaluator - AI Trainer
Business Operations Expert - Evaluator - AI Trainer

Obsidian • Dallas (TX)

On-site
USD 55,000 - 110,000
Information Specialist - Freelance AI Trainer Project
Information Specialist - Freelance AI Trainer Project

Meridial • United States

On-site
Freelance Economics Expert - AI Trainer
Freelance Economics Expert - AI Trainer

Mindrift • Iowa (LA)

Remote
USD 55,000 - 101,000
Competitive hourly rates up to $73
Work flexibility around other commitments
Experience with advanced AI projects
+1
Prompt Engineer
Prompt Engineer

Venteon • Detroit (MI)

On-site
USD 80,000 - 100,000
Medical insurance
Vision insurance
401(k)
Freelance Economics Expert - AI Trainer
Freelance Economics Expert - AI Trainer

Mindrift • Minnesota

Remote
Competitive rates up to $73/hour
Flexible scheduling
Remote work
Remote AI Data Scientist | Prompt Engineering & QA
Remote AI Data Scientist | Prompt Engineering & QA

YO AI Labs • San Francisco (CA)

Remote
USD 55,000 - 96,000