Forward Deployed ML Engineer: Benchmarks & Evaluations

Protege

United States

Remote

USD 120,000 - 180,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Protege is building a platform to accelerate access to AI training data and is hiring a Forward Deployed Machine Learning Engineer for its Benchmarks and Evaluations vertical. You will work closely with the GM and researchers to scale Protege’s leadership in this space and help build the technical foundation for critical eval capabilities.

You will own data pipelines, execution environments, and benchmarking standards, collaborating with researchers to define strong evals across modalities and

Qualifications

  • 4+ years of engineering experience.
  • Hands-on ML work evaluating models.
  • Have previously owned backend and infrastructure.
  • High ambiguity tolerance and bias to action.
  • Comfort working with urgency to meet the pace and volume of the market demands.
  • Strong written communication.

Responsibilities

  • Work on the eval foundation by defining strong evals in different domains and building benchmarks.
  • Own backend infrastructure including data pipelines, execution environments, storage, and orchestration.
  • Go from fast iteration to product by identifying repeatable eval patterns and engaging with DataLab on domain-specific data and research questions.

Skills

Engineering experience
Hands-on ML evaluation
Backend ownership
Ambiguity tolerance
Bias to action
Written communication

Job description

Protege is building a platform to accelerate access to AI training data and is hiring a Forward Deployed Machine Learning Engineer for its Benchmarks and Evaluations vertical. You will work closely with the GM and researchers to scale Protege’s leadership in this space and help build the technical foundation for critical eval capabilities.

You will own data pipelines, execution environments, and benchmarking standards, collaborating with researchers to define strong evals across modalities and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Forward-Deployed ML Engineer — Benchmark & Eval
Remote Forward-Deployed ML Engineer — Benchmark & Eval

Advatix • Northern (KY)

Hybrid
USD 170,000 - 270,000
Forward Deployed ML Engineer
Forward Deployed ML Engineer

Alexander Chapman • United States

On-site
USD 140,000 - 190,000
Remote ML Engineer: Build Benchmarks & Production Pipelines
Remote ML Engineer: Build Benchmarks & Production Pipelines

raydar • Northern (KY)

Hybrid
USD 170,000 - 270,000
Competitive equity
Evaluation Platform Engineer: Build Scalable ML Benchmarks
Evaluation Platform Engineer: Build Scalable ML Benchmarks

Thinking Machines Lab • San Francisco (CA)

On-site
USD 300,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Frontier AI Benchmarks Engineer — Field‑Facing Data & Eval
Frontier AI Benchmarks Engineer — Field‑Facing Data & Eval

Alexander Chapman • United States

On-site
USD 140,000 - 190,000
RL & Data Benchmarking Researcher
RL & Data Benchmarking Researcher

protege • United States

Remote
USD 150,000 - 230,000
Remote Senior AI/ML Benchmarking & Evaluation Engineer
Remote Senior AI/ML Benchmarking & Evaluation Engineer

OpenTeams • Northern (KY)

Hybrid
USD 145,000 - 250,000
Remote AI Benchmark & Datasets Engineer
Remote AI Benchmark & Datasets Engineer

Ignite Next GmbH • Palo Alto (CA), Northern (KY)

Hybrid
USD 120,000 - 190,000
Remote work
Office visits Palo Alto, Paris, Wroclw
Founding Forward Deployed Engineer — New AI Vertical
Founding Forward Deployed Engineer — New AI Vertical

Protege • New City (NY)

On-site
USD 120,000 - 150,000
Forward Deployed Machine Learning Engineer
Forward Deployed Machine Learning Engineer

Advatix • Northern (KY)

Hybrid
USD 170,000 - 270,000