Turn this role into an interview — a resume and cover letter built around what this employer wants.
Agency seeks independent Audio Evaluation Specialists for an AI benchmark project evaluating agentic audio models in real-world customer support scenarios across travel, finance, and telecom. You will design evaluation tasks, simulate interactions, and audit AI outputs to build clean datasets for model refinement.
We require native/bilingual English proficiency and strong communication, plus basic JSON literacy.
We are sourcing independent Audio Evaluation Specialists for an AI benchmark evaluation project assessing advanced agentic audio models. As AI models increasingly handle complex workflows in this domain - specifically real-world customer support scenarios like flight bookings, financial services, and telecommunications, their accuracy relies entirely on robust, expert-crafted training data. The objective of this project is to autonomously produce high-quality evaluation tasks through simulated interactions, audit conversational AI outputs, and generate clean, reliable datasets to optimize model performance.
Operate autonomously to design complex evaluation frameworks and provide structured training data.
To successfully fulfill the deliverables of this project, Contractors must possess deep industry knowledge to craft realistic professional scenarios.
Access to a high-quality microphone to ensure clean, reliable audio input during voice evaluations.
We offer a pay range of $6-to-$65 per hour, with the exact rate determined after evaluating your experience, expertise, and geographic location. Final offer amounts may vary from the pay range listed above.
As a contractor you'll supply a secure computer and high-speed internet; company-sponsored benefits such as health insurance and PTO do not apply.
Engagement Type: Freelance / Independent Contractor
Workplace Type: Remote