Remote Math Research Scientist for AI Evaluation

Crossing Hurdles

United States

Remoto

USD 120.000 - 180.000

Tempo pieno

4 giorni fa
Candidati tra i primi
Generatore di candidature

Non inviare un curriculum generico — genera un curriculum e una lettera di presentazione personalizzati per questo specifico impiego.

Supera i filtri ATS

Descrizione del lavoro

Crossing Hurdles is seeking a researcher to design and solve challenging mathematics problems that test the capabilities and limitations of large language models. You will create clear, accurate, high-quality step-by-step solutions with detailed reasoning and develop gold-standard data for AI model training and evaluation, collaborating with LLM researchers to align tasks with evaluation goals.

The role emphasizes tackling problems from advanced undergraduate to PhD-level concepts in a

Competenze

  • PhD in Mathematics or an equivalent technical field.
  • Strong foundation in advanced mathematics and reasoning.
  • Ability to explain complex concepts with clear, structured step-by-step reasoning.

Mansioni

  • Design and solve challenging mathematics problems testing LLM capabilities.
  • Create clear, accurate, high-quality step-by-step solutions with detailed reasoning.
  • Develop gold-standard mathematical data for AI model training and evaluation.
  • Collaborate with LLM researchers to align task designs with evaluation goals.
  • Target areas where AI models struggle, including symbolic manipulation, abstraction, and multi-step reasoning.
  • Help define evaluation benchmarks covering mathematics from undergraduate to PhD levels.
  • Construct novel mathematical scenarios and evaluate complex reasoning pathways.

Conoscenze

Analytical thinking
Mathematical reasoning
Problem-solving
Structured written communication
English comprehension
Independent work
Self-motivation
Creative thinking
Lateral thinking
Research skills

Formazione

PhD in Mathematics or equivalent technical field

Descrizione del lavoro

Crossing Hurdles is seeking a researcher to design and solve challenging mathematics problems that test the capabilities and limitations of large language models. You will create clear, accurate, high-quality step-by-step solutions with detailed reasoning and develop gold-standard data for AI model training and evaluation, collaborating with LLM researchers to align tasks with evaluation goals.

The role emphasizes tackling problems from advanced undergraduate to PhD-level concepts in a

Ottieni la revisione del curriculum gratis e riservata.

o trascina qui il file.

Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Remote Math Expert for AI Model Benchmarking
Remote Math Expert for AI Model Benchmarking

Parailabs • Northern (KY)

Ibrido
USD 90.000 - 130.000
Fully remote environment
Cutting-edge AI projects
Contractor engagement with potential续
Remote Math Researcher for AI Problem Design (Part-Time)
Remote Math Researcher for AI Problem Design (Part-Time)

Anyone AI Inc. • Stati Uniti

Remoto
USD 45.000 - 65.000
Remote Physics Evaluator & Benchmark Designer
Remote Physics Evaluator & Benchmark Designer

Crossing Hurdles • Stati Uniti

Remoto
USD 55.000 - 83.000
Remote AI Math Research Specialist for LLM Training
Remote AI Math Research Specialist for LLM Training

Weekday AI • Stati Uniti

Remoto
USD 1.240.000 - 2.204.000
Remote AI Math Problem Designer & Model Evaluator
Remote AI Math Problem Designer & Model Evaluator

RemoExperts • Stati Uniti

Remoto
USD 103.000 - 172.000
Maths Expert
Maths Expert

Crossing Hurdles • Stati Uniti

Remoto
USD 120.000 - 180.000
Remote AI Evaluation Scientist — Advanced STEM Problems
Remote AI Evaluation Scientist — Advanced STEM Problems

24-Mag Llc • New York (NY)

Remoto
USD 80.000 - 150.000
Remote Math AI Researcher (PhD) - Contract (20+ hrs/wk)
Remote Math AI Researcher (PhD) - Contract (20+ hrs/wk)

Turing • New York (NY)

Remoto
USD 41.328 - 82.656
Remote Math Research Collaborator for AI Reasoning
Remote Math Research Collaborator for AI Reasoning

Weekday AI • Stati Uniti

Remoto
USD 110.000 - 152.000
Fully remote
Remote AI Evaluation & Computational Math Specialist
Remote AI Evaluation & Computational Math Specialist

OpenTrain AI, Inc. • Stati Uniti

Remoto
USD 90.000 - 120.000