Stand out for this role — generate a tailored resume and cover letter in about a minute.
Obsidian is building high-fidelity simulated work environments to evaluate AI agents in real marketing tasks. You will design content marketing scenarios, define required documents, dashboards, personas, and tool states for realistic evaluation.
You will write and review reference answers and scoring criteria, and assess AI agent attempts against a clear standard. The role spans diagnosing data, prioritizing, planning, QA, and cross-tool orchestration across marketing stack.
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Content Marketing environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Research and discovery, editorial prioritization, content production, and content QA and audit.
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.