Remote ML Evaluation Architect (Research-Driven)

Alignerr Corp.

Greater London

Remote

GBP 55,000 - 138,000

Part time

42 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Alignerr is seeking researchers with deep machine learning knowledge to design evaluation challenges that push the limits of today’s AI systems. This is an hourly remote contract offering 10–40 hours per week, with autonomy over your schedule and collaboration with a global team.

You will develop complex problems grounded in your domain expertise, craft rigorous evaluation criteria, and contribute to gold-standard solutions that shape how models are measured and improved.

Qualifications

  • Deep ML knowledge and ability to design evaluation problems.
  • Ability to identify gaps where general ML knowledge falls short.
  • Experience publishing original research is valued.
  • Excellent written communication; articulate complex problems clearly.

Responsibilities

  • Design complex, original machine learning problems rooted in your domain.
  • Craft evaluation tasks that require advanced domain knowledge beyond standard ML pipelines.
  • Draw from your research to create problems that challenge state-of-the-art AI.
  • Define rigorous problem statements, evaluation criteria, and gold-standard solutions.
  • Assess AI-generated ML solutions for correctness, creativity, and methodological soundness.
  • Document problem difficulty levels, required domain knowledge, and AI failure modes.
  • Collaborate asynchronously with a global team of researchers and engineers.

Skills

ML knowledge
Domain expertise
Research experience
Strong writing
Independent work

Education

MS or PhD

Job description

Alignerr is seeking researchers with deep machine learning knowledge to design evaluation challenges that push the limits of today’s AI systems. This is an hourly remote contract offering 10–40 hours per week, with autonomy over your schedule and collaboration with a global team.

You will develop complex problems grounded in your domain expertise, craft rigorous evaluation criteria, and contribute to gold-standard solutions that shape how models are measured and improved.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Machine Learning Evaluation Specialist
Machine Learning Evaluation Specialist

Alignerr Corp. • Greater London

Remote
GBP 55,000 - 138,000
Remote ML Engineer for AI Training & Code Review
Remote ML Engineer for AI Training & Code Review

Alignerr Corp. • Greater London

On-site
GBP 7,577,000 - 13,087,000
Software Engineer – Machine Learning (AI Training)
Software Engineer – Machine Learning (AI Training)

Alignerr Corp. • Greater London

On-site
GBP 7,577,000 - 13,087,000
AI Research Engineer: Evaluation & Alignment Scientist
AI Research Engineer: Evaluation & Alignment Scientist

Prolific • United Kingdom

Remote
GBP 95,000 - 130,000
Remote work
Competitive benefits
Remote AI Data Trainer - Senior Data Scientist (Contract)
Remote AI Data Trainer - Senior Data Scientist (Contract)

Alignerr Corp. • Greater London

Remote
GBP 83,000 - 124,000
AI/ML Engineer — RLHF & Model Evaluation (Remote)
AI/ML Engineer — RLHF & Model Evaluation (Remote)

Prolific • Sheffield

On-site
GBP 67,249 - 95,779
AI/ML Engineer — RLHF & Model Evaluation (Remote)
AI/ML Engineer — RLHF & Model Evaluation (Remote)

Prolific • Glasgow

On-site
GBP 67,249 - 95,779
Senior Machine Learning Expert
Senior Machine Learning Expert

Alignerr Corp. • Greater London

Remote
GBP 96,000 - 165,000
Remote Senior ML Expert: Reasoning & Trace Architect
Remote Senior ML Expert: Reasoning & Trace Architect

Alignerr Corp. • Greater London

Remote
GBP 96,000 - 165,000
AI/ML Engineer — RLHF & Model Evaluation (Remote)
AI/ML Engineer — RLHF & Model Evaluation (Remote)

Prolific • Greater London

On-site
GBP 67,299 - 95,850