Machine Learning Evaluation Specialist

Alignerr Corp.

Greater London

Remote

GBP 55,000 - 138,000

Part time

33 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Alignerr is seeking researchers with deep machine learning knowledge to design evaluation challenges that push the limits of today’s AI systems. This is an hourly remote contract offering 10–40 hours per week, with autonomy over your schedule and collaboration with a global team.

You will develop complex problems grounded in your domain expertise, craft rigorous evaluation criteria, and contribute to gold-standard solutions that shape how models are measured and improved.

Qualifications

  • Deep ML knowledge and ability to design evaluation problems.
  • Ability to identify gaps where general ML knowledge falls short.
  • Experience publishing original research is valued.
  • Excellent written communication; articulate complex problems clearly.

Responsibilities

  • Design complex, original machine learning problems rooted in your domain.
  • Craft evaluation tasks that require advanced domain knowledge beyond standard ML pipelines.
  • Draw from your research to create problems that challenge state-of-the-art AI.
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions.
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness.
  • Document problem difficulty levels, required domain knowledge, and AI failure modes.
  • Collaborate asynchronously with a global team of researchers and engineers.

Skills

ML knowledge
Domain expertise
Research experience
Strong writing
Independent work

Education

MS or PhD

Job description

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define — and push — the limits of today's most capable AI systems.

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define — and push — the limits of today's most capable AI systems.

This isn't routine review work. You'll apply your hard-earned research expertise to craft problems that state-of-the-art models genuinely struggle to solve. Your contributions directly shape how the next generation of AI is measured, benchmarked, and improved.

  • Organization: Alignerr
  • Type: Hourly Contract
  • Location: Fully Remote
  • Commitment: 10–40 hours/week

What You'll Do

  • Design complex, original machine learning problems rooted in your specific domain of expertise
  • Craft evaluation tasks that require advanced domain knowledge well beyond standard ML pipelines
  • Draw from your own research experience to create problems that genuinely challenge state-of-the-art AI
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness
  • Document problem difficulty levels, required domain knowledge, and expected AI failure modes
  • Collaborate asynchronously with a global team of researchers and engineers

Who You Are

  • Graduate-level expertise (MS or PhD preferred) in a scientific or technical domain that intersects with machine learning
  • Strong working knowledge of ML methods — model selection, feature engineering, evaluation metrics, and pipeline design
  • Deep familiarity with active research problems in your field
  • Able to identify precisely where general ML knowledge falls short and specialized domain insight becomes critical
  • Experience publishing or conducting original research is highly valued
  • Excellent written communication — you can articulate complex problems clearly and precisely
  • Self-motivated and comfortable working independently on intellectually demanding tasks

Example Domains (Not Exhaustive)

  • Computational biology, genomics, or bioinformatics
  • Climate science and environmental modeling
  • Medical imaging and healthcare ML
  • Materials science and computational chemistry
  • Astrophysics and signal processing
  • Natural language processing for low-resource or specialized corpora
  • Robotics, control theory, or reinforcement learning in complex environments
  • Financial modeling and quantitative analysis

Why Join Us

  • Work at the true frontier of AI evaluation and safety research
  • Collaborate with top research labs pushing the boundaries of what AI can do
  • Finally put your specialized domain expertise to use in a high-impact, meaningful way
  • Full autonomy over your schedule — work when and how you do your best thinking
  • Flexible, fully remote contract with potential for ongoing work and deeper research involvement
  • Build your profile as a recognized contributor to cutting-edge AI development
  • Join a global community of researchers and engineers who take this work seriously

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Machine Learning Expert
Senior Machine Learning Expert

Alignerr Corp. • Greater London

Remote
GBP 96,000 - 165,000
Software Engineer – Machine Learning (AI Training)
Software Engineer – Machine Learning (AI Training)

Alignerr Corp. • Greater London

On-site
GBP 7,577,000 - 13,087,000
Remote ML Evaluation Architect (Research-Driven)
Remote ML Evaluation Architect (Research-Driven)

Alignerr Corp. • Greater London

Remote
GBP 55,000 - 138,000
Data Scientist (Masters)
Data Scientist (Masters)

Alignerr Corp. • Greater London

Remote
GBP 83,000 - 124,000
AI Native Builder
AI Native Builder

Newpage Solutions • Bristol

On-site
GBP 70,000 - 110,000
Flexible remote-first collaboration
People-first culture
Global, collaborative team
Software Engineer (AI Training)
Software Engineer (AI Training)

Alignerr Corp. • Manchester

Remote
GBP 34,000 - 62,000
AI Scientist
AI Scientist

Seven Sigma Group • Greater London

Remote
GBP 100,000 - 150,000
Top-of-market compensation
Ownership of AI products from idea to‑
Direct access to latest AI tools and A
+2
Applied AI/ML Engineer for Production Deployments
Applied AI/ML Engineer for Production Deployments

brainco • Greater London

On-site
GBP 90,000 - 130,000
Competitive salary
Paid maternity and paternity leave
Daily lunches
+2
Machine Learning Engineer (Contract)
Machine Learning Engineer (Contract)

AND Digital Limited • Milton Keynes

On-site
GBP 60,000 - 110,000
Public Policy Economist (AI Training)
Public Policy Economist (AI Training)

Alignerr Corp. • City of Edinburgh

Remote
GBP 55,000 - 110,000
Autonomy
Variety
Global collaboration
+2