Verschicke keinen generischen Lebenslauf — erstelle einen Lebenslauf und ein Anschreiben, die genau auf diese Rolle zugeschnitten sind.
Obsidian is building high-fidelity simulated work environments to evaluate AI agents in marketing contexts. You design task briefs, specify required documents and tools, and craft reference answers and scoring rubrics to differentiate strong from weak work.
You will review AI agent attempts, judge against a rigorous standard, and handle cross-tool orchestration across tasks such as diagnosis, prioritization, planning, QA and reporting.
We are building high-fidelity simulated work environments used to evaluate and improve AI agents on real marketing work. Each environment reproduces a marketing org's actual tool surface — email, storage, CRM, project management, social media management, web analytics, AEO/SEO, ads, CMS, product analytics and support — populated with realistic documents, dashboards, personas and deliberately planted problems.
You will help design and pressure-test the Creator and Influencer environments: the briefs, the artifacts, the judgment calls a strong practitioner would make, and the errors a weaker one would miss.
Tasks span the capabilities we measure: diagnosing what happened from messy or conflicting data, prioritizing and making tradeoffs, planning and executing, QA and reconciliation, triage and escalation, research and evaluation, reporting, and orchestrating multi-step work across several tools.
Creator discovery and evaluation, influencer program decisions, seeding and campaign execution, usage rights and disclosure judgment, and share-of-voice contribution.
Applicants complete a short multiple-choice knowledge screener specific to this sub-domain before review.