Remote STEM & Technical Expert: Rubrics & AI Evaluation

AI Trainer Jobs

United States

Remote

USD 114,000 - 162,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

AI Trainer Jobs is seeking a STEM & Technical SME to craft expert-level questions, evaluate AI outputs, and design rigorous rubrics in a remote hourly contractor role.

You will develop problems that require multi-step reasoning, identify weaknesses in frontier AI models, and translate complex judgments into clear, measurable criteria, including partial-credit standards and edge-case considerations. You will contribute to RLHF/SFT workflows while maintaining high professional standards in English.

Qualifications

  • Bachelor’s degree or higher in Mathematics, Physics, Chemistry, Engineering, Computer Science, Statistics, Applied Science, or another technical discipline.
  • Strong professional proficiency in English, minimum C1, with the ability to write precise technical explanations and evaluation criteria.
  • 3+ years of professional, academic, research, or industry experience in your stated technical domain.
  • Demonstrated advanced expertise in a clearly defined technical field or specialization.
  • Ability to create challenging technical questions that require multi-step reasoning, domain expertise, or professional judgment.
  • Strong ability to design detailed evaluation rubrics defining required reasoning, correct methodology, acceptable alternatives, partial-credit criteria, and critical errors.
  • Ability to distinguish between a correct final answer and an answer supported by valid versus flawed reasoning.
  • Comfortable reviewing and critiquing other experts’ rubrics for ambiguity, missing criteria, redundancy, technical inaccuracies, or poor scoring design.
  • High attention to detail when evaluating formulas, assumptions, units, methodology, edge cases, logical consistency, and technical terminology.
  • Experience with exam writing, academic grading, peer review, research review, technical QA, standards development, or assessment design is strongly preferred.
  • Prior experience with AI evaluation, RLHF, SFT, benchmarking, data annotation, prompt design, or LLM evaluation is preferred.
  • Reliable, self-directed, and able to deliver consistent quality in an hourly, remote contractor workflow.

Responsibilities

  • Create expert-level technical questions: Develop difficult questions that test genuine technical reasoning and expose weaknesses in advanced AI systems.
  • Design evaluation rubrics: Write structured criteria defining what a correct, rigorous, and complete response must contain.
  • Define partial-credit standards: Identify essential reasoning steps, acceptable alternatives, minor errors, and critical failures.
  • Evaluate AI-generated solutions: Assess outputs for correctness, reasoning quality, completeness, methodology, and technical precision.
  • Identify hidden reasoning errors: Detect cases where an answer reaches the correct conclusion using invalid assumptions or flawed logic.
  • Critique peer rubrics: Review other experts’ criteria for technical accuracy, clarity, completeness, and scoring consistency.
  • Develop reference content: Produce expert answers, explanations, critiques, and other gold-standard content.
  • Support AI training: Create questions, rubrics, evaluations, preference judgments, and reference-answer pairs suitable for RL and SFT workflows.
  • Maintain technical rigor: Ensure content reflects appropriate professional or academic standards within the relevant discipline.

Skills

STEM
Mathematics
Physics
Chemistry
Engineering
Computer Science
Technical Reasoning
AI Evaluation
Rubric Design
Rubric Writing
Benchmarking
Model Evaluation
Question Writing
Expert Review
English

Education

Bachelor’s degree or higher in a technical discipline

Job description

AI Trainer Jobs is seeking a STEM & Technical SME to craft expert-level questions, evaluate AI outputs, and design rigorous rubrics in a remote hourly contractor role.

You will develop problems that require multi-step reasoning, identify weaknesses in frontier AI models, and translate complex judgments into clear, measurable criteria, including partial-credit standards and edge-case considerations. You will contribute to RLHF/SFT workflows while maintaining high professional standards in English.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

STEM & Technical Specialist – Rubric Design
STEM & Technical Specialist – Rubric Design

AI Trainer Jobs • United States

Remote
USD 114,000 - 162,000
STEM Rubric Design Specialist
STEM Rubric Design Specialist

OpenTrain AI, Inc. • Northern (KY)

Remote
USD 117,000 - 158,000
Remote work
Part-time contractor role
AI Evaluation Rubric Designer
AI Evaluation Rubric Designer

OpenTrain AI, Inc. • Northern (KY)

Remote
USD 117,000 - 158,000
Remote work
Part-time contractor role
Remote Rubrics QA Expert — AI Workflow Evaluator
Remote Rubrics QA Expert — AI Workflow Evaluator

AI Trainer Jobs • United States

Remote
USD 48,000 - 69,000
Remote JSON Rubrics Expert — AI Workflow QA
Remote JSON Rubrics Expert — AI Workflow QA

AI Trainer Jobs • United States

Remote
USD 69,000 - 241,000
Remote AI Evaluation Rubric Reviewer | Labeling & Calibration
Remote AI Evaluation Rubric Reviewer | Labeling & Calibration

AI Trainer Jobs • United States

Remote
USD 28,000 - 62,000
Scientific and Technical Service Specialist - Freelance AI Trainer Project
Scientific and Technical Service Specialist - Freelance AI Trainer Project

Invisible Technologies Inc. • United States

Remote
USD 14,000 - 41,000
Remote HR Operations AI Evaluator & Rubric Reviewer
Remote HR Operations AI Evaluator & Rubric Reviewer

AI Trainer Jobs • United States

Remote
USD 28,000 - 55,000
Remote JS/TS AI Evaluator & Rubric Specialist
Remote JS/TS AI Evaluator & Rubric Specialist

AI Trainer Jobs • United States

Remote
USD 34,000 - 55,000
Remote AI Content Evaluator & Rubric Specialist
Remote AI Content Evaluator & Rubric Specialist

24-Mag Llc • Northern (KY), New York (NY)

Hybrid
USD 124,000 - 193,000