Get more replies from employers
Send a job-specific resume in minutes.
Obsidian in New York builds high‑fidelity simulated work environments to evaluate AI agents on real marketing tasks. You will design creator and influencer scenarios, including briefs, artifacts, and the judging criteria that separate strong responses from plausible but incorrect ones.
Your work will specify required documents, dashboards, personas, and tool states, write or review reference answers, and evaluate AI attempts against your rubric.
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Creator and Influencer environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Creator discovery and evaluation, influencer program decisions, seeding and campaign execution, usage rights and disclosure judgment, and share-of-voice contribution.
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.