Health & Medicine Benchmark Architect — Remote

Appsierra Group

American Samoa

On-site

USD 129,000 - 164,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Appsierra Group is seeking a remote Health & Medicine Benchmark Specialist to develop and review rigorous academic assessment content for an AI research initiative. You will author and verify questions in health and medicine, evaluate solution quality, and contribute to high-quality benchmarks used to advance AI capabilities.

Responsibilities include question authoring or verification, rating difficulty, and providing detailed reasoning and references.

Qualifications

  • MD/DO/PhD or doctoral candidate in Medicine, Biomedical Sciences, Public Health, or closely related field.
  • Master's degree may be considered for exceptional expertise in a subdomain.
  • Strong graduate-level knowledge of medicine and health sciences.
  • Strong clinical reasoning and biomedical research methodology skills.
  • Board certification, clinical experience, or health-related publications advantageous.
  • Excellent written English with ability to communicate complex concepts clearly and concisely.

Responsibilities

  • Create original, challenging questions that assess deep conceptual understanding.
  • Ensure each question is unambiguous, self-contained, precisely defined, and includes all information necessary to solve it.
  • Rate question difficulty: Medium = Undergraduate, Hard = Advanced Undergraduate, Expert = Postgraduate level.
  • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives.
  • Develop clear, concise step-by-step reasoning in markdown format.
  • Provide 1–5 reputable academic references per question.

Skills

Clinical reasoning
Biomedical research methods
Academic writing

Education

MD/DO/PhD or doctoral candidate in Medicine/Biomedical Sciences/Public Health
Master's degree considered for exceptional candidates

Job description

Appsierra Group is seeking a remote Health & Medicine Benchmark Specialist to develop and review rigorous academic assessment content for an AI research initiative. You will author and verify questions in health and medicine, evaluate solution quality, and contribute to high-quality benchmarks used to advance AI capabilities.

Responsibilities include question authoring or verification, rating difficulty, and providing detailed reasoning and references.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Health & Medicine AI Benchmark Specialist
Remote Health & Medicine AI Benchmark Specialist

Appsierra Group • United States

On-site
USD 129,000 - 164,000
Remote Health AI Benchmark Architect
Remote Health AI Benchmark Architect

Mercor • United States

Remote
USD 55,000 - 110,000
Medical AI Assessment Architect — Remote Content & Review
Medical AI Assessment Architect — Remote Content & Review

Obsidian • Nashville (TN)

On-site
USD 26,000 - 52,000
Remote Medical Assessment Content Architect
Remote Medical Assessment Content Architect

Obsidian • New York (NY)

On-site
USD 70,000 - 110,000
Fully remote
Flexible schedule
AI Benchmark Engineer — PhD (Remote & Flexible)
AI Benchmark Engineer — PhD (Remote & Flexible)

Mercor • New York (NY)

On-site
USD 55,000 - 110,000
Biostatistician for AI Benchmarking — Remote Contractor
Biostatistician for AI Benchmarking — Remote Contractor

Appsierra Group • American Samoa

On-site
USD 125,000 - 208,000
Remote work
Remote Medical Evaluation Specialist for AI Benchmarking
Remote Medical Evaluation Specialist for AI Benchmarking

YO AI Labs • Boston (MA)

Remote
USD 55,000 - 110,000
Remote Medical Evaluation Architect for AI Benchmarks
Remote Medical Evaluation Architect for AI Benchmarks

YO AI Labs • Dallas (TX)

Remote
USD 83,000 - 152,000
Remote AI Benchmark Engineer — PhD-Level Question Author
Remote AI Benchmark Engineer — PhD-Level Question Author

Obsidian • Detroit (MI)

On-site
USD 90,000 - 130,000
Remote Medical Evaluation Specialist for AI Benchmarking
Remote Medical Evaluation Specialist for AI Benchmarking

YO AI Labs • San Jose (CA)

Remote
USD 83,000 - 138,000