Turn this role into an interview — a resume and cover letter built around what this employer wants.
Morpheus Talent Solutions seeks a founding Applied AI Researcher focused on model evaluation and data strategy for an early-stage, profitable AI research company. This research-first seat requires designing evaluations, testing hypotheses about model failures, and defining data schemas, rubrics, and quality controls in collaboration with engineers to scale programs.
You will publish studies that position the company as a research partner to frontier AI labs, contributing to the field and
Applied AI Researcher - Model Evaluation & Data Strategy
San Francisco (in-person preferred; open to remote across US, UK, Australia, and Europe) · Retained search - confidential client
Morpheus has been exclusively retained to lead the search for a founding Applied AI Researcher on behalf of an early-stage, profitable AI research company.
An early-stage AI research company that works with leading AI labs to find where frontier models fail and build the expert human data that fixes them. They run a vetted network of 5,000+ top-1% specialists across finance, medicine, law, engineering, music, and other domains - competing on the quality of expert judgment, not the scale of cheap labeling.
Backed by a top pre-seed fund and angel investors who are founders and senior researchers at leading frontier AI labs. Already profitable.
A founding, research-first seat - a genuine thought partner on evaluation and data strategy, not someone who coordinates other people's research, and not client-facing or delivery. You'll design evaluations, form and test hypotheses about model failure, and define the datasets, rubrics, reward signals, and quality controls that move performance - then work with engineers to turn them into scalable programs. Publishing is core to this role, not a perk - you'll be expected to author and present research that positions the company as a research partner to the field, so a prior publication record is essential.
Experience at an AI lab, foundation-model company, or post-training team; expert-data or human-eval program design; multimodal/coding/agentic eval work; public benchmarks or eval frameworks.
This is an early, high-momentum team that currently works a six-day week: Saturdays are fully remote and self-directed, no set hours - most people use them as a heads-down research day. Compensation is $200K-$350K base + equity; visa sponsorship available.