AI Research Peer Review Evaluator (ML/AI)

lightly

Zürich

Remote

CHF 83.000 - 124.000

Teilzeit

14 Tage+
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die genau auf das zugeschnitten sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Fully remote
Part-time
Flexible hours

Zusammenfassung

Lightly AG in Zurich, an ETH/HSG spin-off, is seeking researchers with strong ML/AI backgrounds to support an AI evaluation project focused on scientific peer review.

You’ll evaluate reviews generated by agentic AI systems and compare them against expert human peer reviews of ML/AI research papers. This is a remote, project-based contractor opportunity with flexible working hours.

Qualifikationen

  • Master’s degree or PhD in ML/AI, CS, or a closely related field.
  • Experience reading ML/AI research papers, evaluating methodology and claims.
  • Strong written communication and ability to provide concise rationales.

Aufgaben

  • Read and scan ML/AI research papers to understand core contributions, methods, experiments, and claims.
  • Review original human peer reviews to establish a benchmark for each paper.
  • Evaluate AI-generated reviews against the baseline using a structured rubric.
  • Assess accuracy, depth, value, and novelty of each AI review.
  • Identify hallucinations, unsupported claims, or overlooked insights.
  • Compare two AI reviews side-by-side and determine stronger analyses.
  • Search and verify relevant literature from Google Scholar, arXiv, etc.
  • Provide concise, evidence-based rationales and apply evaluation guidelines consistently.

Kenntnisse

ML/AI
Peer review
Analytical writing
Critical reading papers
Literature search

Ausbildung

Master’s or PhD in ML/AI/CS

Tools

Google Scholar
arXiv
Semantic Scholar

Jobbeschreibung

Lightly AG is a Zurich-based AI company and ETH/HSG spin-off, backed by Y Combinator and top-tier investors. Our machine learning and computer vision technology is trusted by global leaders in autonomous driving, medical imaging, and visual inspection.

We’re looking for researchers with strong Machine Learning / AI backgrounds to support an AI evaluation project focused on scientific peer review. You’ll evaluate reviews generated by agentic AI systems and compare them against expert human peer reviews of ML/AI research papers.

This is a remote, project-based contractor opportunity with flexible working hours.

Tasks

What you'll be doing

  • Read and scan ML/AI research papers to understand their core contributions, methodology, experiments, and claims
  • Review the original human peer reviews to establish an expert baseline for each paper
  • Evaluate AI-generated peer reviews against that baseline using a structured scoring rubric
  • Assess the technical accuracy, analytical depth, constructive value, and novelty/significance assessment of each AI review
  • Identify hallucinations, unsupported claims, missed technical issues, or valuable insights surfaced by the AI reviewers
  • Compare two AI-generated reviews side-by-side and determine where one provides stronger or more useful analysis
  • Search and verify relevant academic literature using sources such as Google Scholar, arXiv, or Semantic Scholar, including checking whether cited prior work was available before the paper’s submission date
  • Provide concise, evidence-based rationales explaining your evaluation decisions and consistently apply the project rubric

The evaluation specifically looks at whether agentic AI reviewers can provide meaningful value beyond expert human reviewers—for example, by identifying relevant prior literature that humans missed, questioning important assumptions, or resolving inconsistencies using evidence.

Requirements

You're a strong candidate if you:

  • Have a Master’s, PhD, or are currently pursuing graduate study in Machine Learning, Artificial Intelligence, Computer Science, Statistics, or a closely related technical field
  • Have contributed to at least one scientific/research paper, ideally as a first author, although co-authors and other substantial contributors are also welcome
  • Have experience critically reading ML/AI research papers, including evaluating methodology, experimental design, results, limitations, and scientific claims
  • Are familiar with major ML/AI research venues, such as NeurIPS, ICML, ICLR, ACL, CVPR, or comparable conferences and journals
  • Have prior academic peer-review experience, ideally for an ML/AI conference or journal — strongly preferred
  • Are comfortable conducting academic literature searches and verifying prior work, publication dates, citations, and novelty claims
  • Have strong analytical and written communication skills and can distinguish meaningful technical concerns from superficial criticism
  • Can provide clear, concise, evidence-based rationales for your decisions
  • Can consistently apply detailed evaluation guidelines and scoring rubrics across multiple papers and reviews
  • Have strong attention to detail, particularly when identifying factual inaccuracies or hallucinated technical claims
Benefits
  • Fully remote and flexible — work from anywhere
  • Part-time contractor role with flexible hours
  • Work directly on the evaluation of cutting-edge agentic AI systems for scientific research
  • Apply your ML/AI research expertise to help measure and improve the quality of AI-generated scientific peer review

We look forward to hearing from you!

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Research Peer Review Evaluator (ML/AI)
AI Research Peer Review Evaluator (ML/AI)

Join • Zürich

Vor Ort
CHF 83.000 - 138.000
Fully remote
Part-time contract
AI Research Peer Review Evaluator (ML/AI)
AI Research Peer Review Evaluator (ML/AI)

your Jared • Zürich

Hybrid
CHF 34.000 - 57.000
Remote ML/AI Peer Review Evaluator
Remote ML/AI Peer Review Evaluator

Join • Zürich

Hybrid
CHF 83.000 - 138.000
Fully remote
Part-time contract
Remote ML/AI Research Peer Review Specialist
Remote ML/AI Research Peer Review Specialist

your Jared • Zürich

Hybrid
CHF 34.000 - 57.000
Journalist/Writer
Journalist/Writer

RemoteJobsOne • Zürich

Remote
CHF 103.000 - 160.000
Senior AI Engineer (f/m/x)
Senior AI Engineer (f/m/x)

Lever, Inc. • Schweiz

Remote
CHF 103.000 - 150.000
Remote-first
Vienna office
Flexible hours
+7
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Interiman Group • Basel

Vor Ort
CHF 150.000 - 230.000
AI Scientist - LLM Systems
AI Scientist - LLM Systems

Artificialy • Lugano

Vor Ort
CHF 120.000 - 180.000
Competitive compensation
Growth opportunities
Scientific environment
+1
AI Scientist - LLM Systems
AI Scientist - LLM Systems

Artificialy • Zürich

Vor Ort
CHF 130.000 - 180.000
Researcher - Agentic Algorithms and Architectures
Researcher - Agentic Algorithms and Architectures

Huawei Switzerland • Zürich

Vor Ort
CHF 120.000 - 180.000
Competitive salary
Growth opportunities
Innovative projects
+1