ML Engineer - AI Coding Expert

Mercor

New York (NY)

On-site

USD 220,000 - 551,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Mercor is seeking contributors to evaluate frontier AI coding models and support a Frontier Code Agents project from its base in New York. You will review model-generated ML components, focusing on training, inference, MLOps, and LLM applications, while spotting bugs and performance gaps.

The role emphasizes hands-on evaluation, professional judgment, and working with production ML workflows on a sprint-based schedule. Prior ML deployment experience is valued.

Qualifications

  • 2+ years of professional ML engineering experience.
  • Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated ML implementations and technical tradeoffs.
  • Experience deploying ML systems to production is preferred.

Responsibilities

  • Use frontier AI coding agents to complete and evaluate complex ML and AI engineering tasks.
  • Review model-generated implementations involving training, inference, MLOps, or LLM apps.
  • Identify bugs, edge cases, performance issues, and failure modes.
  • Compare outputs from frontier models and assess strengths and weaknesses.
  • Apply professional engineering judgment to realistic ML engineering scenarios.

Skills

ML engineering experience
Production ML systems
AI coding agents usage
Evaluate model-generated ML
ML deployment experience

Tools

Cursor
Claude Code
Codex
Windsurf
Gemini CLI

Job description

About the Role
  • Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project.
  • Contributors help evaluate and improve frontier AI coding models through structured technical assessments.
  • The work focuses on realistic machine learning engineering workflows and model evaluation.
  • Spots are limited and filling quickly on a first come, first serve basis.
What You'll Do
  • Use frontier AI coding agents to complete and evaluate complex machine learning and AI engineering tasks.
  • Review model-generated implementations involving model training, inference systems, MLOps, and LLM applications.
  • Identify bugs, edge cases, performance issues, and failure modes.
  • Compare outputs from multiple frontier models and assess their strengths and weaknesses.
  • Apply professional engineering judgment to realistic ML engineering scenarios.
Time Commitment
  • Sprint based project that runs in 12-24 hour stretches based on client requirement.
Compensation
  • $400 per accepted task.
  • Typical tasks take approximately 2–3 hours after ramp-up.
  • Compensation is tied to accepted work.
Who Should Apply
  • 2+ years of professional machine learning engineering experience.
  • Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated machine learning implementations and technical tradeoffs.
  • Experience deploying ML systems to production is preferred.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Engineer - Coding Agent Expert
ML Engineer - Coding Agent Expert

Obsidian • New York (NY)

Hybrid
ML Engineer - AI Coding Expert - AI Trainer
ML Engineer - AI Coding Expert - AI Trainer

Mercor • Miami (FL)

On-site
USD 455,000 - 647,000
Frontier AI ML Engineer — Real-World Model Evaluator
Frontier AI ML Engineer — Real-World Model Evaluator

Great Value Hiring • United States

On-site
USD 55,104 - 82,656
Data Engineer - AI Coding Expert
Data Engineer - AI Coding Expert

Obsidian • New York (NY)

On-site
USD 454,608 - 647,472
ML Engineer
ML Engineer

Great Value Hiring • United States

On-site
USD 55,104 - 82,656
Data Engineer - AI Model Evaluator - AI Trainer
Data Engineer - AI Model Evaluator - AI Trainer

Obsidian • Philadelphia

On-site
USD 455,000 - 647,000
Data Engineer - AI Model Evaluator - AI Trainer
Data Engineer - AI Model Evaluator - AI Trainer

Mercor • Philadelphia

On-site
USD 179,000 - 276,000
Data Engineer - AI Model Evaluator
Data Engineer - AI Model Evaluator

Mercor • New York (NY)

On-site
USD 38,000 - 46,000
Data Engineer - AI Model Evaluation
Data Engineer - AI Model Evaluation

Obsidian • San Francisco (CA)

On-site
USD 207,000 - 551,000
Fraud Engineer - AI Specialist
Fraud Engineer - AI Specialist

Obsidian • San Francisco (CA)

On-site