Machine Learning Evaluation Specialist

Alignerr Corp.

Greater London

On-site

GBP 83,000 - 165,000

Full time

10 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Alignerr is seeking researchers and domain experts to design evaluation challenges that push the boundaries of AI systems. You will craft complex ML problems rooted in your domain knowledge and define rigorous evaluation criteria, including gold-standard solutions.

You will collaborate asynchronously with a global team, publish or build on existing research, and enjoy full autonomy over your schedule in a flexible, remote hourly contract.

Qualifications

  • Graduate-level expertise in a scientific or technical domain that intersects with machine learning.
  • Strong working knowledge of ML methods - model selection, feature engineering, evaluation metrics, and pipeline design.
  • Deep familiarity with active research problems in your field.
  • Able to identify precisely where general ML knowledge falls short and specialized domain insight becomes critical.
  • Experience publishing or conducting original research is highly valued.
  • Excellent written communication - you can articulate complex problems clearly and precisely.
  • Self-motivated and comfortable working independently on intellectually demanding tasks.

Responsibilities

  • Design complex, original machine learning problems rooted in your specific domain of expertise
  • Craft evaluation tasks that require advanced domain knowledge well beyond standard ML pipelines
  • Draw from your own research experience to create problems that genuinely challenge state-of-the-art AI
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness
  • Document problem difficulty levels, required domain knowledge, and expected AI failure modes
  • Collaborate asynchronously with a global team of researchers and engineers

Skills

ML knowledge
Model selection
Feature engineering
Evaluation metrics
ML pipeline design
Written communication
Independent work

Education

MS or PhD preferred

Job description

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define — and push — the limits of today's most capable AI systems.

About The Role

The quality of AI depends entirely on the quality of the problems used to test it. We're looking for researchers and domain experts with deep machine learning knowledge to design the evaluation challenges that define — and push — the limits of today's most capable AI systems.

  • Organization: Alignerr
  • Type: Hourly Contract
  • Location: Fully Remote
  • Commitment: 10-40 hours/week
What You'll Do
  • Design complex, original machine learning problems rooted in your specific domain of expertise
  • Craft evaluation tasks that require advanced domain knowledge well beyond standard ML pipelines
  • Draw from your own research experience to create problems that genuinely challenge state-of-the-art AI
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness
  • Document problem difficulty levels, required domain knowledge, and expected AI failure modes
  • Collaborate asynchronously with a global team of researchers and engineers
Who You Are
  • Graduate-level expertise (MS or PhD preferred) in a scientific or technical domain that intersects with machine learning
  • Strong working knowledge of ML methods - model selection, feature engineering, evaluation metrics, and pipeline design
  • Deep familiarity with active research problems in your field
  • Able to identify precisely where general ML knowledge falls short and specialized domain insight becomes critical
  • Experience publishing or conducting original research is highly valued
  • Excellent written communication - you can articulate complex problems clearly and precisely
  • Self-motivated and comfortable working independently on intellectually demanding tasks
Example Domains (Not Exhaustive)
  • Computational biology, genomics, or bioinformatics
  • Climate science and environmental modeling
  • Medical imaging and healthcare ML
  • Materials science and computational chemistry
  • Astrophysics and signal processing
  • Natural language processing for low-resource or specialized corpora
  • Robotics, control theory, or reinforcement learning in complex environments
  • Financial modeling and quantitative analysis
Why Join Us
  • Work at the true frontier of AI evaluation and safety research
  • Collaborate with top research labs pushing the boundaries of what AI can do
  • Finally put your specialized domain expertise to use in a high-impact, meaningful way
  • Full autonomy over your schedule - work when and how you do your best thinking
  • Flexible, fully remote contract with potential for ongoing work and deeper research involvement
  • Build your profile as a recognized contributor to cutting-edge AI development
  • Join a global community of researchers and engineers who take this work seriously
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer – Machine Learning (AI Training)
Software Engineer – Machine Learning (AI Training)

Alignerr Corp. • Greater London

On-site
GBP 7,577,000 - 13,087,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Alignerr Corp. • Greater London

Remote
GBP 62,000 - 125,000
Electrical Engineering
Electrical Engineering

Alignerr Corp. • Birmingham

Remote
GBP 55,000 - 117,000
Fully remote
Flexible schedule
Global collaboration
Remote ML Evaluation Scientist - Domain Expert
Remote ML Evaluation Scientist - Domain Expert

Alignerr Corp. • Greater London

On-site
GBP 83,000 - 165,000
Mechanical Engineering
Mechanical Engineering

Alignerr Corp. • Cambridge

Remote
GBP 62,000 - 125,000
Remote freelance work
Flexible schedule
Incident Response Analyst
Incident Response Analyst

Alignerr Corp. • Manchester

On-site
GBP 55,000 - 96,000
ML Engineer
ML Engineer

RemoteJobsOne • Greater London

Remote
GBP 83,000 - 156,000
Data Science Expert - AI Evaluation
Data Science Expert - AI Evaluation

Mercor • Greater London

On-site
GBP 90,000 - 130,000
AI Research Engineer (Remote - UK)
AI Research Engineer (Remote - UK)

Jobgether • United Kingdom

On-site
GBP 60,000 - 80,000
Competitive salary
Flexible remote work arrangements
Professional growth opportunities
Machine Learning Engineer, Platform
Machine Learning Engineer, Platform

Mat Vin • Greater London

On-site
GBP 90,000 - 130,000