Frontier AI Code Engineer

Mercor

New York (NY)

On-site

USD 220,000 - 551,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Mercor is seeking contributors to evaluate frontier AI coding models and support a Frontier Code Agents project from its base in New York. You will review model-generated ML components, focusing on training, inference, MLOps, and LLM applications, while spotting bugs and performance gaps.

The role emphasizes hands-on evaluation, professional judgment, and working with production ML workflows on a sprint-based schedule. Prior ML deployment experience is valued.

Qualifications

  • 2+ years of professional ML engineering experience.
  • Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products.
  • Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools.
  • Ability to evaluate model-generated ML implementations and technical tradeoffs.
  • Experience deploying ML systems to production is preferred.

Responsibilities

  • Use frontier AI coding agents to complete and evaluate complex ML and AI engineering tasks.
  • Review model-generated implementations involving training, inference, MLOps, or LLM apps.
  • Identify bugs, edge cases, performance issues, and failure modes.
  • Compare outputs from frontier models and assess strengths and weaknesses.
  • Apply professional engineering judgment to realistic ML engineering scenarios.

Skills

ML engineering experience
Production ML systems
AI coding agents usage
Evaluate model-generated ML
ML deployment experience

Tools

Cursor
Claude Code
Codex
Windsurf
Gemini CLI

Job description

Mercor is seeking contributors to evaluate frontier AI coding models and support a Frontier Code Agents project from its base in New York. You will review model-generated ML components, focusing on training, inference, MLOps, and LLM applications, while spotting bugs and performance gaps.

The role emphasizes hands-on evaluation, professional judgment, and working with production ML workflows on a sprint-based schedule. Prior ML deployment experience is valued.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Engineer: Frontier AI Coding Evaluator
ML Engineer: Frontier AI Coding Evaluator

Mercor • Miami (FL)

On-site
USD 455,000 - 647,000
Frontier ML Engineer: Benchmark AI Coding Models
Frontier ML Engineer: Benchmark AI Coding Models

Obsidian • New York (NY)

On-site
USD 207,000 - 303,000
ML Engineer - AI Coding Expert
ML Engineer - AI Coding Expert

Mercor • New York (NY)

On-site
USD 220,000 - 551,000
ML Engineer - Coding Agent Expert
ML Engineer - Coding Agent Expert

Obsidian • New York (NY)

Hybrid
AI-Driven DevOps Engineer & Model Evaluator
AI-Driven DevOps Engineer & Model Evaluator

Mercor • San Francisco (CA)

On-site
USD 15,000 - 22,000
ML Engineer - AI Coding Expert - AI Trainer
ML Engineer - AI Coding Expert - AI Trainer

Mercor • Miami (FL)

On-site
USD 455,000 - 647,000
Frontier AI ML Engineer — Evaluation & Production Systems
Frontier AI ML Engineer — Evaluation & Production Systems

Obsidian • New York (NY)

Remote
USD 4,000 - 7,000
Frontier ML Engineer: AI Coding Agent Evaluator
Frontier ML Engineer: AI Coding Agent Evaluator

Obsidian • New York (NY)

Hybrid
Frontier AI Data Engineer: ETL & Model Evaluation
Frontier AI Data Engineer: ETL & Model Evaluation

Mercor • Philadelphia

On-site
USD 179,000 - 276,000
Frontier AI Data Engineer: Model Evaluation & Pipelines
Frontier AI Data Engineer: Model Evaluation & Pipelines

Obsidian • Philadelphia

On-site
USD 455,000 - 647,000