A complete application in a minute — tailored resume and cover letter, ready to send.
Mercor in New York seeks a senior researcher with deep expertise in control systems, analog/RF circuits, power electronics, or mechanical design to lead a frontier engineering-reasoning evaluation with an AI research lab. The role involves defining simulations, specs, and prompts, and guiding the model to generate designs that satisfy every specification.
You will assess the model's internal reasoning on your tasks, using an agentic grader to score spec satisfaction and ensure rigorous
A frontier engineering-reasoning evaluation run in collaboration with a leading AI research lab. The work measures whether state-of-the‑art models can reason from first principles in your engineering domain rather than retrieve facts from training data — and you get visibility into the model's internal reasoning on your own tasks.
Domain experts write simulations and specification sheets; the model attempts to design an artifact — controller gains, circuit parameters, geometry — that satisfies every spec.
You define the simulation, a set of specs with pass/fail thresholds, and the prompt. The model probes your simulation with a limited number of calls, then submits a final design. An agentic grader runs the simulation and scores spec satisfaction.
Onboarding calls run daily, alongside internal tooling built to help you work faster. Prior model-evaluation experience is not required.