A complete application in a minute — tailored resume and cover letter, ready to send.
Salesforce is seeking an Associate Evaluations Manager for the Digital Success Data & AI team. You will measure, monitor, and improve Agentforce performance on Salesforce Help, with emphasis on quality, latency, and instruction adherence.
You will build scalable evaluation frameworks, collaborate with Support Delivery, Engineering, Operations, and Data Science, and translate results into defensible insights guiding product decisions and launches.
Program & Project Management
About Salesforce
Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn't a buzzword - it's a way of life. The world of work as we know it is changing and we're looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce's core values at the heart of it all.
Ready to level-up your career at the company leading workforce transformation in the agentic era? You're in the right place! Agentforce is the future of AI, and you are the future of Salesforce.
As an Associate Evaluations Manager on the Digital Success Data & AI team, you will help measure, monitor, and improve the performance of Agentforce on Salesforce Help. You will execute ongoing synthetic evaluations across Answer Quality, latency, instruction adherence, and other core capabilities, translating results into clear, actionable insights that inform product decisions, operational improvements, and leadership reporting.
You will contribute to operational excellence by building clear documentation, repeatable processes, and scalable evaluation frameworks that increase confidence in agent performance. This role partners closely with Support Delivery, Engineering, Operations, and Data Science to evaluate new features, identify risks and opportunities, and ensure quality at launch.
You will also help evolve our evaluation ecosystem by refining LLM judge prompts, golden datasets, and evaluation assets while identifying opportunities to increase automation, reliability, and efficiency. Success in this role requires analytical rigor, curiosity, and ownership-someone comfortable working through ambiguity, solving problems end-to-end, and consistently delivering defensible insights that drive action.