Get more replies from employers
Send a job-specific resume in minutes.
United States Digital Space LLC seeks an AI Evaluation Engineer to join our AI Evaluation team in Canada. You will own evaluation coverage for the company’s Agentic AI systems alongside the evaluation lead, focusing on LLM-judge metrics, scenario and benchmark dataset curation, and error analysis to support release-readiness for voice and chat solutions.
This role reports to the AI Evaluation manager and may be based in our Vancouver office.
the company is the AI platform for customer experience, built to resolve customer problems in real time across voice and digital. Our AI agents learn from your best human agents and improve with every interaction, helping organizations understand their customers, deliver better experiences, increase operational efficiencies, and build a lasting competitive advantage.
Unlike legacy systems built to route and answer, or standalone agentic bot vendors built to deflect, the company was built to resolve. Our AI agents and human agents operate on a single platform with shared context, allowing Agentic AI to resolve issues, advance deals, and eliminate busywork through automation while seamlessly handing conversations to humans when needed, with full context preserved.
Market-leading brands, including Randstad, Motorola Solutions, Netflix, the San Diego Padres, the Colorado Rockies Baseball Club, and Cal Athletics, trust the company. the company is backed by Andreessen Horowitz, GV, ICONIQ Capital, and T-Mobile.
At the company, AI isn’t just a feature; it’s how our teams do their best work every day. We put powerful AI tools in every employee’s hands so they can move faster, think bigger, and achieve more.
We believe every conversation matters. And we’ve built the platform that turns those conversations into insight and action, for our customers and ourselves.
We look for people who are intensely curious and hold themselves to a high bar. Our ambition is significant, and achieving it requires a team that operates at the highest level. We seek individuals who embody our core traits: Scrappy, Curious, Optimistic, Persistent, and Empathetic.
As an AI Evaluation Engineer, you'll be an integral part of our AI Evaluation team, owning evaluation coverage for the company’s Agentic AI systems alongside our existing evaluation lead. A key focus will be co-owning LLM-judge metric development and calibration, scenario and benchmark dataset curation, and structured error analysis to support release-readiness decisions for our agentic voice and chat solutions.
This position reports to the manager of the AI Evaluation team and has the opportunity to be based in our Vancouver office.
For exceptional talent based in Ontario, Canadathe target base salary range for this position is posted below. Our salary ranges are determined by role, level, and location. The range displayed on each job posting reflects the