An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Mercor is building high-fidelity simulated work environments to evaluate and improve AI agents on real marketing work. The team designs lifecycle and CRM scenarios, specifying the required documents, dashboards, personas and tool states, and writing reference answers with clear criteria for strong performance.
The role focuses on diagnosing issues from messy data, prioritizing tradeoffs, planning, QA, and cross-tool coordination, with emphasis on explainable reasoning and robust evaluation of
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Lifecycle environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Contact strategy, milestone campaigns, customer communications, deliverability and list health, and lifecycle portfolio optimization.
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.