AI Benchmarking Lead, Performance Benchmarking Evaluation

Amazon Inc.

Hyderabad

On-site

INR 2,000,000 - 3,600,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Amazon is seeking an AI Benchmarking Lead in Hyderabad to ensure reliable evaluation metrics for Seller Assistant. The role focuses on auditing processes, calibration, and quality assurance across ML data workflows, with cross-team coordination and SOP enhancements.

The candidate will drive continuous improvement in audit methodologies, surface issues early, and collaborate with program managers and applied scientists to scale seller coverage and performance globally.

Qualifications

  • Experience in natural language data labeling or linguistic annotation.
  • Proficiency with MS Excel; basic SQL and Python understanding.
  • Strong English communication skills (verbal and written).

Responsibilities

  • Evaluate audits by core auditing team to improve evaluation metrics.
  • Improve audit reliability and consistency through systematic measurement.
  • Calibrate processes to ensure quality standards across audits.
  • Quality-check audits and provide actionable feedback to the team.
  • Drive continuous improvement in audit processes and methodologies.
  • Perform quality checks on audits performed by the team.
  • Identify rubric gaps and ambiguities leading to inconsistent outcomes.
  • Surface high-confidence product issues by validating model failures.
  • Coordinate annotation tasks across ML data processes to ensure quality.
  • Understand dependencies across ML data workflows and communicate impact.
  • Update and document SOP changes; secure approvals and train the team.
  • Test new SOPs/tools and provide improvement feedback for onboarding.

Skills

Data labeling experience
MS Excel
English communication
SQL basics
Python basics

Education

Bachelor's degree or equivalent in a related field

Job description

Job ID: 10563414 | ADCI HYD 13 SEZ - H84

Join our mission-critical team supporting Seller Assistant, Amazon's Gen-AI powered copilot that helps sellers navigate Amazon's complex ecosystem and grow their businesses. As a AI Benchmarking Lead, you'll play a pivotal role in ensuring the reliability and accuracy of AI model evaluations as we scale from 61% to 90%+ active seller coverage worldwide.

About Seller Assistant

Seller Assistant is a conversational AI copilot that understands the full context of a seller's business. It intelligently orchestrates back end tools to deliver actionable, drilled-down responses and can independently complete complex tasks on behalf of sellers with their permission.

Our Scale and Impact
  • Expanded to 2.44MM sellers (45x growth vs. Dec 2024)
  • Currently serving 61% of active sellers worldwide across 9 international stores (CN2XX, IN, UK, DE, JP, BR, MX, AE, SA)
  • Supporting four languages: English, Chinese, German, and Japanese
  • 2026 Goal: Scale to 90%+ active sellers WW with 5 new store launches (France, Italy, Spain, Canada, Australia)
AI Benchmarking Lead Responsibilities
  • Evaluate audits performed by the core auditing team to increase confidence in evaluation metrics
  • Improve audit reliability and consistency through systematic measurement of auditor accuracy
  • Conduct targeted calibration to ensure quality standards across the auditing function
  • Enforce quality standards by quality-checking audits and providing actionable feedback to team members
  • Drive continuous improvement in audit processes and methodologies
Additional Responsibilities
  • You conduct quality checks on audits performed by the core auditing team.
  • You identify rubric gaps and evaluation ambiguities that lead to inconsistent audit outcomes.
  • You surface high-confidence product issues earlier by validating and categorizing model failures.
  • You serve as point of contact for annotation tasks across ML data process areas, ensuring quality execution and delivery.
  • You understand dependencies across ML data workflows and articulate customer impact effectively.
  • You modify existing annotation methods and update SOPs.
  • You document SOP changes, secure approval, share knowledge with the team, and audit adoption and execution.
  • You test new SOPs and tools, providing feedback on quality and improvement recommendations to support onboarding.
Key job responsibilities
  • You structure data collection, analyse results and share inputs for SOP changes.
  • You collate, track, and report progress on key metrics agreed to with respective stakeholders (e.g., Program managers, Applied Scientist) specific to your functional area.
  • You identify operational issues related to process and tooling and recommend suggestions to improve key project metrics such as productivity and quality.
Basic Qualifications
  • Bachelor's degree or equivalent in a related field
  • Experience in natural language data labeling, data annotation, linguistic annotation or other forms of data markup
  • Technical Skills: Proficiency in MS Excel; basic understanding of SQL and Python
  • Experience with Microsoft Office products and applicationsCommunication Skills: Strong verbal and written communication skills in English
  • Knowledge about SOA and process that deal with sellers.
Preferred Qualifications
  • 1 to 3 years of equivalent experience
  • Performed annotation related tasks across ML data process areas.
  • Strong knowledge of process documentation, analysis knowledge
  • Technical proficiency in SQL querying and Python programming for data analysis
  • Strong analytical and problem-solving skills
  • Ability to work independently and as part of a team

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status. Veterans, military spouses, and people with disabilities are encouraged to apply.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Benchmarking Specialist Performance Benchmarking Evaluation
AI Benchmarking Specialist Performance Benchmarking Evaluation

Amazon • Hyderabad

On-site
INR 350,000 - 520,000
AIT Audit Snr Manager, Audits and Insights Team
AIT Audit Snr Manager, Audits and Insights Team

Amazon Inc. • Hyderabad

On-site
INR 5,500,000 - 7,500,000
Business Intel Engineer, GO-AI
Business Intel Engineer, GO-AI

Amazon • Hyderabad

On-site
INR 1,800,000 - 2,400,000
ML Data Associate-II, Artificial General Intelligence Data Services
ML Data Associate-II, Artificial General Intelligence Data Services

Amazon Inc. • Chennai District

On-site
INR 500,000 - 750,000
Associate II, ML Data Operations, GO-AI Operations
Associate II, ML Data Operations, GO-AI Operations

Amazon • Hyderabad

On-site
INR 320,000 - 520,000
Software Development Engineer III, Seller and AM GenAI Tools
Software Development Engineer III, Seller and AM GenAI Tools

Amazon Inc. • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Business Intel Engineer, GO-AI
Business Intel Engineer, GO-AI

Amazon Inc. • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Team Manager, RBKS AI Data
Team Manager, RBKS AI Data

Amazon Inc. • Hyderabad

On-site
INR 800,000 - 1,200,000
Software Development Engineer III, Seller and AM GenAI Tools
Software Development Engineer III, Seller and AM GenAI Tools

Amazon • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Software Development Engineer, Seller and AM GenAI Tools
Software Development Engineer, Seller and AM GenAI Tools

Amazon • Hyderabad

On-site
INR 1,500,000 - 2,100,000