Turn this role into an interview — a resume and cover letter built around what this employer wants.
Obsidian is building high-fidelity simulated field marketing environments used to evaluate and improve AI agents on real marketing work. You will design task briefs, specify the artifacts and tool states needed, and craft the reference answers and evaluation criteria.
In this role you will review AI agent attempts, judge performance, and support the design of exercises that test diagnostic reasoning, prioritization, planning, and cross-tool orchestration across CRM, analytics and project
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Field Marketing environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Event follow-up, sponsorship decisions, regional pipeline recovery, and joint planning against a named account list.
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.