ML Engineer: Frontier AI Coding Evaluator

Mercor

Miami (FL)

On-site

USD 455,000 - 647,000

Part time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments.

The work focuses on realistic machine learning engineering workflows and model evaluation. As a contractor you will use frontier AI coding agents to complete and evaluate complex ML engineering tasks, review model-generated implementations, identify bugs and performance issues, and compare outputs

Qualifications

  • 2+ years of professional machine learning engineering experience.
  • Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.

Responsibilities

  • Use frontier AI coding agents to complete and evaluate complex machine learning and AI engineering tasks.
  • Review model-generated implementations involving model training, inference systems, MLOps, and LLM applications.
  • Identify bugs, edge cases, performance issues, and failure modes.
  • Compare outputs from multiple frontier models and assess their strengths and weaknesses.
  • Apply professional engineering judgment to realistic ML engineering scenarios.

Skills

ML engineering
Production ML systems
MLOps
LLM applications
AI coding agents
Model deployment

Tools

Cursor
Claude Code
Codex
Windsurf
Gemini CLI

Job description

Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. Contributors help evaluate and improve frontier AI coding models through structured technical assessments.

The work focuses on realistic machine learning engineering workflows and model evaluation. As a contractor you will use frontier AI coding agents to complete and evaluate complex ML engineering tasks, review model-generated implementations, identify bugs and performance issues, and compare outputs

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Frontier AI Code Engineer
Frontier AI Code Engineer

Mercor • New York (NY)

On-site
USD 220,000 - 551,000
ML Engineer - AI Coding Expert - AI Trainer
ML Engineer - AI Coding Expert - AI Trainer

Mercor • Miami (FL)

On-site
USD 455,000 - 647,000
Frontier AI Data Engineer: ETL & Model Evaluation
Frontier AI Data Engineer: ETL & Model Evaluation

Mercor • Philadelphia

On-site
USD 179,000 - 276,000
Frontier AI Infrastructure Engineer (Contract)
Frontier AI Infrastructure Engineer (Contract)

Mercor • Miami (FL)

On-site
USD 165,000 - 276,000
Frontier AI Data Engineer: Model Evaluation & Pipelines
Frontier AI Data Engineer: Model Evaluation & Pipelines

Obsidian • Philadelphia

On-site
USD 455,000 - 647,000
AI-Powered Data Engineer & Model Evaluator
AI-Powered Data Engineer & Model Evaluator

Mercor • New York (NY)

On-site
USD 38,000 - 46,000
Frontier ML Engineer: AI Coding Agent Evaluator
Frontier ML Engineer: AI Coding Agent Evaluator

Obsidian • New York (NY)

Hybrid
USD 275,520 - 826,560
ML Engineer: Frontier AI Code Agent Evaluator
ML Engineer: Frontier AI Code Agent Evaluator

Obsidian • Chicago (IL)

On-site
USD 454,608 - 647,472
AI-Driven DevOps Engineer & Model Evaluator
AI-Driven DevOps Engineer & Model Evaluator

Mercor • San Francisco (CA)

On-site
USD 15,000 - 22,000
ML Engineer - AI Coding Expert
ML Engineer - AI Coding Expert

Mercor • New York (NY)

On-site
USD 220,000 - 551,000