Research Engineer, Environments

Mercor

San Francisco (CA)

On-site

USD 150,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Semi-annual bonus
Generous equity
Relocation bonus up to $15k
Housing stipend up to $10k
Meals stipend $1.5k monthly
Equinox membership
Laundry reimbursement
Wellness reimbursement
Health, Dental, Vision insurance

Job summary

Mercor seeks engineers to capture enterprise data and transform it into high-fidelity RL environments for evaluation and dataset training. You will automate the process of building evals for real work in the economy and push the frontier of world-building and verifier engineering.

You will collaborate with researchers, operators, and AI partners, contributing to scalable, high-quality simulated environments used by Fortune 500 companies.

Qualifications

  • Experience shipping environments or contributing to OSS frameworks.
  • Experience building environments at previous companies or working on agentic evaluations.
  • Strong full-stack engineering skills spanning infrastructure to app code.

Responsibilities

  • Ship models for workflow extraction, classification, and grading.
  • Engineer autonomous task refinement processes to distill data into pipelines.
  • Deliver data to customers and deploy into real engagements.
  • Help define the future of agentic transformation for enterprises worldwide.
  • Build end-to-end environments for labs and enterprises by platformizing sandbox apps and loading real data.

Skills

Full-stack engineering
OSS/environments
Systems thinking
Data/ RL environments

Tools

Temporal
Modal
Orchestration services

Job description

About Mercor

About MercorMercor's mission to organize human intelligence to power the AI economy. We're a leading AI data company, building the layer between human expertise and frontier models. Millions of domain experts on the platform are paid over $4 million per day to train frontier AI models. Mercor's APEX benchmark family measures AI's real-world impact on professional work. Mercor Enterprise brings this same infrastructure to Fortune 500 companies: helping companies capture how their best people actually work, translating that expertise directly back into agents. Mercor is creating a new category of work where expertise powers AI advancement. Achieving this requires an ambitious, fast-paced and deeply committed team. You’ll work alongside researchers, operators, and AI companies at the forefront of shaping the systems that are redefining society. Mercor is a profitable Series C company valued at $10 billion. We work in-person five days a week in our San Francisco, NYC, or London offices.


About the Role

You’ll work with large enterprises to capture their data and transform it into high-fidelity RL environments for capability evaluations and training datasets for frontier labs. We focus on pushing the frontier of world-building, verifier engineering, and more alongside our partners.Your goal will be to automate the process of building evals for real work in the economy.


What You\'ll Do


  • Ship models for workflow extraction, classification, and grading.

  • Engineer autonomous task refinement processes which distill data taste into pipelines.

  • Deliver data to customers and deploy into real engagements.

  • Help define the future of agentic transformation for enterprises around the world.

  • Deeply learn about the intricacies of enterprises through building evaluations for all aspects of work.

  • Build end-to-end environments for labs & enterprises by platformizing sandbox app clones, load real data into the sandboxes, build prompts from real workflows, and write verifiers leveraging enterprise expertise & golden outputs.

  • Systematize the production of environments to scale throughput while maintaining high-quality worlds and verifiers.


What We're Looking For


  • Prior experience shipping environments – you\'ve contributed to an OSS framework, built environments at previous companies, or worked on agentic evaluations.

  • Strong full-stack engineering skills – you\'ll be responsible for everything from infrastructure to app code to analytics

  • Bias to action – this team is focused on shipping evals, not just philosophizing about them.

  • Curiosity – being biased towards understanding and digging deep into model behavior and actually looking at the data.

  • Sweat the details that make a simulation indistinguishable from the real thing and have systems-level thinking skills that allow you to scale up quality.


Nice to Have


  • Experience with Temporal, Modal, or similar orchestration/compute services

  • Experience with synthetic data generation for frontier models.

  • Past work auditing and scrutinizing industry-standard evaluations


Benefits


  • Semi-annual performance bonus structure

  • Generous equity grant vested over 4 years

  • Up to $15k Relocation bonus

  • $10K housing bonus (if you live within 0.5 miles of our office)

  • $1.5K monthly stipend for meals

  • Free Equinox membership

  • $200 monthly laundry reimbursement

  • $200 monthly personal wellness reimbursement

  • Health, Dental, Vision insurance


<\/p>
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Enterprise Evals Platform
Member of Technical Staff, Enterprise Evals Platform

Doist • San Francisco (CA)

On-site
USD 150,000 - 190,000
Relocation bonus
Housing stipend
Meals stipend
+5
Research Engineer – Benchmarking
Research Engineer – Benchmarking

Doist • San Francisco (CA)

On-site
USD 150,000 - 210,000
Bi-annual bonus
Equity grant
Relocation bonus
+6
Member of Technical Staff, Enterprise Platform
Member of Technical Staff, Enterprise Platform

Apply • New York (NY)

On-site
USD 170,000 - 260,000
Relocation bonus
Housing bonus
Meal stipend
+7
Member of Technical Staff, Enterprise Platform
Member of Technical Staff, Enterprise Platform

Doist • New York (NY), San Francisco (CA)

On-site
USD 170,000 - 250,000
Up to $15k relocation bonus
$10K housing bonus
$1.5K monthly meals stipend
+5
Deployment Strategist
Deployment Strategist

Mercor • San Francisco (CA)

On-site
USD 140,000 - 210,000
Bi-annual performance bonus
Generous equity grant
Relocation bonus up to $15k
+6
Research Engineer – Benchmarking
Research Engineer – Benchmarking

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
Bi-annual bonus
Equity grant
Relocation bonus
+8
Member of Technical Staff, Deeptune
Member of Technical Staff, Deeptune

Mercor • New York (NY)

On-site
USD 220,000 - 600,000
Relocation bonus
Housing bonus
Meals stipend
+7
Product Engineer, Talent Experience
Product Engineer, Talent Experience

Mercor, Inc. • San Francisco (CA)

On-site
USD 150,000 - 210,000
Bi-annual performance bonus
Generous equity grant
Relocation bonus up to $15k
+4
Research Engineer – Benchmarking
Research Engineer – Benchmarking

Mercor • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Bi-annual performance bonus
Generous equity grant
Up to $15k relocation bonus
+5
Software Engineer, Backend
Software Engineer, Backend

Mercor • San Francisco (CA)

On-site
USD 180,000 - 300,000
Bi-annual performance bonus
Generous equity grant
Relocation bonus
+8