Remote AI Benchmarking Consultant (Part-Time)

Obsidian

San Francisco (CA)

On-site

USD 27,000 - 55,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job description

Location requirements

Canada United States

Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Management Consulting domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.

We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.

Contract and Payment Terms
  • You will be engaged as an independent contractor.
  • This is a fully remote role that can be completed on your own schedule.
  • Projects can be extended, shortened, or concluded early depending on needs and performance.
  • Your work at Mercor will not involve access to confidential or proprietary information from any employer, client, or institution.
  • Payments are weekly on Stripe or Wise based on services rendered.
  • Please note: We are unable to support H1-B or STEM OPT candidates at this time.
About Mercor

Mercor partners with leading AI labs and enterprises to train frontier models using human expertise. You will work on projects that focus on training and enhancing AI systems. You will be paid competitively, collaborate with leading researchers, and help shape the next generation of AI systems in your area of expertise.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Management Consultant - Document Specialist
Management Consultant - Document Specialist

Obsidian • San Francisco (CA)

Remote
USD 27,000 - 55,000
Remote Underwriting AI Benchmark Specialist
Remote Underwriting AI Benchmark Specialist

Obsidian • San Francisco (CA)

On-site
USD 1,775,000 - 2,730,000
Remote AI Benchmarking Consultant (Part-Time)
Remote AI Benchmarking Consultant (Part-Time)

Mercor • San Francisco (CA)

On-site
Underwriting Expert - Document Analysis
Underwriting Expert - Document Analysis

Obsidian • San Francisco (CA)

Remote
USD 1,775,000 - 2,730,000
Remote Applied Sciences AI Benchmark Specialist
Remote Applied Sciences AI Benchmark Specialist

Mercor • San Francisco (CA)

On-site
Remote Educational AI Benchmark Researcher (Part-Time)
Remote Educational AI Benchmark Researcher (Part-Time)

Mercor • San Francisco (CA)

Remote
USD 83,000 - 124,000
Remote Education AI Benchmark Specialist (Contractor)
Remote Education AI Benchmark Specialist (Contractor)

Mercor • New York (NY)

Remote
USD 14,000 - 20,000
Remote AI Document Benchmark Specialist
Remote AI Document Benchmark Specialist

Mercor • New York (NY)

Remote
USD 25,000 - 50,000
AI Education Benchmark Specialist (Remote)
AI Education Benchmark Specialist (Remote)

Obsidian • San Francisco (CA)

On-site
USD 27,000 - 46,000
Remote Underwriting Expert — AI Benchmarking Projects
Remote Underwriting Expert — AI Benchmarking Projects

Mercor • San Francisco (CA)

On-site
USD 60,000 - 80,000