GenAI Retail Evaluator for LLM Training

DigiNo

Northern (KY)

Hybrid

USD 140,000 - 190,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Cincinnatus LLC is seeking a seasoned Retail SME to join its GenAI initiative and help evaluate AI model outputs against real-world retail rubrics. You will support research and engineering teams to bring retail judgment to AI training data and ensure high-quality evaluations.

Ideal candidates have 8+ years in retail, proven ability to work 35 hours weekly, and strong communication and problem-solving skills. This is a W-2 engagement with potential placement at a leading AI Lab.

Qualifications

  • 8+ years of professional experience in retail (merchandising, category management, operations, buying/planning) at top-tier organizations.
  • Hands-on experience evaluating LLM/AI model outputs against rubrics or structured scoring criteria.
  • Demonstrable career progression (e.g., Category Manager, Senior Manager, Director).

Responsibilities

  • Guide research and engineering teams to close knowledge gaps in retail merchandising, category management, and operations reasoning.
  • Design challenging retail tasks and write accurate, well-reasoned solutions grounded in real retail practice.
  • Evaluate AI model outputs against rubrics and provide clear, written feedback on correctness, judgment, and reasoning quality.
  • Develop and refine evaluation guidelines and scoring rubrics specific to retail tasks.
  • Collaborate with other SMEs to ensure consistency and accuracy in training data.

Skills

Retail experience
AI model evaluation
Career progression
Written communication

Job description

Cincinnatus LLC is seeking a seasoned Retail SME to join its GenAI initiative and help evaluate AI model outputs against real-world retail rubrics. You will support research and engineering teams to bring retail judgment to AI training data and ensure high-quality evaluations.

Ideal candidates have 8+ years in retail, proven ability to work 35 hours weekly, and strong communication and problem-solving skills. This is a W-2 engagement with potential placement at a leading AI Lab.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Retail AI Evaluation Lead
Senior Retail AI Evaluation Lead

Obsidian • New York (NY)

Remote
USD 70,000 - 90,000
Retail AI Evaluation Specialist
Retail AI Evaluation Specialist

Mercor • New York (NY)

Remote
USD 110,000 - 170,000
Senior Retail AI Evaluation Specialist
Senior Retail AI Evaluation Specialist

Mercor • New York (NY)

On-site
USD 120,000 - 180,000
Retail SME - AI Evaluation Expert
Retail SME - AI Evaluation Expert

Mercor • New York (NY)

On-site
USD 120,000 - 180,000
Retail SME - AI Evaluation Expert
Retail SME - AI Evaluation Expert

Obsidian • New York (NY)

On-site
USD 120,000 - 170,000
Retail Specialist
Retail Specialist

Mercor • United States

On-site
USD 120,000 - 180,000
Retail Specialist
Retail Specialist

DigiNo • Northern (KY)

Hybrid
USD 140,000 - 190,000
GenAI Marketing Strategist & Model Evaluator
GenAI Marketing Strategist & Model Evaluator

Mercor • United States

On-site
USD 120,000 - 180,000
GenAI Marketing Strategist — AI Training Evaluation
GenAI Marketing Strategist — AI Training Evaluation

Mercor • New York (NY)

On-site
USD 130,000 - 180,000
GenAI Marketing Strategist for AI Training Data
GenAI Marketing Strategist for AI Training Data

Mercor • San Francisco (CA)

Remote
USD 120,000 - 180,000