Machine Learning Benchmark Task Designer

ixolabs.ai

France

Hybride

EUR 72 000 - 168 000

Plein temps

Il y a 3 jours
Soyez parmi les premiers à postuler
Générateur de candidature

Transformez ce poste en entretien — un CV et une lettre de motivation conçus selon ce que cet employeur recherche.

Passez les filtres ATS

Résumé du poste

IXO is engaging specialists to evaluate machine-learning coding tasks with realistic data and verifiable results. You will construct or assess Python-based modeling and evaluation tasks for AI engineering systems.

Remote assignments are scheduled by agreement, with about 15 hours per week depending on the track. The pay is per completed task, USD, and the fee is agreed before work; no guaranteed volume is promised.

Qualifications

  • A master's or PhD in computing, ML, AI, engineering, mathematics, statistics or a related quantitative subject.
  • Strong ML experience with practical Python coding.
  • Experience with Python-based modeling, data inspection and reproducible checks.

Responsabilités

  • Construct or assess Python-based modeling and evaluation tasks for AI systems.
  • Develop a correct reference approach and inspect data, metrics and assumptions.
  • Provide reproducible checks and clear technical explanations for errors or alternatives.

Connaissances

Python
ML evaluation
Code review

Formation

Master's/PhD in computing/ML/AI/engineering/math/statistics

Outils

JAX
PyTorch
SciPy/NumPy
vLLM
SGLang
llama.cpp
Hugging Face models/tokenizers

Description du poste

IXO is engaging specialists to evaluate machine-learning coding tasks with realistic data and verifiable results. Your contribution is technical work: make the reasoning inspectable, identify substantive errors and provide evidence that supports a reliable assessment.

Work you will do
  • Construct or assess Python-based modeling and evaluation tasks for AI engineering systems.
  • Develop a correct reference approach and inspect the data, metrics and implementation assumptions that govern the result.
  • Provide reproducible checks and clear technical explanations for important errors or alternative solutions.
Required background and routes
  • A master's or PhD in computing, ML, AI, engineering, mathematics, statistics or a related quantitative subject, with strong professional or research ML experience.
  • Practical Python and coding-agent skills, using suitable tools such as JAX, PyTorch, SciPy/NumPy, vLLM, SGLang, llama.cpp or Hugging Face models/tokenizers; equivalent relevant tools are acceptable.
Preferred background
  • Experience in an established engineering or AI research organization. Exceptional academic or open-source work can also qualify.
Deliverables

Submit the completed technical artifact or assessment with its supporting evidence, explicit assumptions, reproducible checks where applicable, and concise reasons for each material judgment. Address review findings within the agreed scope.

Location and schedule

Remote assignments are scheduled by agreement, with no guaranteed weekly volume. Availability planning can include 15 hours per week depending on the track. IXO confirms the applicable timing before work.

Pay and working terms

Per completed task; fee agreed before work USD. Payment is for a completed task that meets the agreed acceptance criteria. The fee, task specification and revision requirements are agreed before work; an advertised time estimate does not determine payment. Applying does not guarantee an assignment. Use public, licensed or otherwise authorized material only; do not submit confidential employer information, personal data or restricted research.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Published ML Benchmark Researcher
Published ML Benchmark Researcher

ixolabs.ai • France

Hybride
EUR 180 000 - 192 000
ML Experiment Validity Reviewer
ML Experiment Validity Reviewer

ixolabs.ai • France

Hybride
EUR 90 000 - 114 000
Command-Line Engineering Environment Author
Command-Line Engineering Environment Author

ixolabs.ai • France

Hybride
EUR 52 000 - 78 000
Remote ML Benchmark Task Architect
Remote ML Benchmark Task Architect

ixolabs.ai • France

Sur place
EUR 72 000 - 168 000
Mechanical Simulation Task Designer
Mechanical Simulation Task Designer

ixolabs.ai • France

Hybride
EUR 72 000 - 144 000
MCP Software Evaluation Environment Author
MCP Software Evaluation Environment Author

ixolabs.ai • France

Hybride
EUR 7 800 - 22 000
Remote work opportunities
Flexible scheduling
Remote ML Benchmark Evaluator & Task Designer
Remote ML Benchmark Evaluator & Task Designer

ixolabs.ai • France

Sur place
EUR 180 000 - 192 000
Materials ML Research Benchmark Specialist
Materials ML Research Benchmark Specialist

ixolabs.ai • France

Hybride
EUR 192 000 - 251 000
Full-Stack Software Benchmark Task Author
Full-Stack Software Benchmark Task Author

ixolabs.ai • France

Hybride
Multidiscipline Engineering Challenge Task Author
Multidiscipline Engineering Challenge Task Author

ixolabs.ai • France

Hybride
EUR 49 000 - 116 000