A complete application in a minute — tailored resume and cover letter, ready to send.
ByteDance in Dubai is offering a Project Intern position within the AI Evaluation team. You will work on defining and executing evaluation benchmarks for AI agents, including creating datasets and rubrics, and analyzing results to drive improvements.
The role involves collaborating with the AI evaluation team across global offices, leveraging frontline tools and internal systems, and gaining hands-on experience in a fast-moving tech environment.
Join us as we work together to inspire creativity and enrich life around the globe.
Location:
Dubai
Team:
Product
Employment Type:
Intern
Job Code:
A62230B
Share this listing:
We are the AI Vertical Scenarios Team within ByteDance's Corporate Service System, based in Dubai. Our mission is to set the standard for AI native applications across ByteDance's internal workplace and procurement scenarios, and to drive consistent AI experiences for employees globally. We build flagship AI applications in-house and use a full scope evaluation suite as our measurement foundation to track how well leading models perform on the real, day to day tasks ByteDance employees carry out.
Why join us
As a Project Intern, you will contribute to impactful short-term projects and gain hands‑on experience in a fast‑paced, professional environment. This internship offers the opportunity to develop practical skills, apply your knowledge to real‑world challenges, and explore your career interests.
Applications are reviewed on a rolling basis, so we encourage you to apply early.
Online AssessmentCandidates who pass resume screening will be invited to participate in Our Company's technical online assessment.
Core evaluation work: Own a defined scope of ByteDance's AI evaluation benchmark end to end, with ownership to identify problems and propose solutions. You'll run the full research and evaluation loop on some of the world's most widely used AI Agents, surfacing capability gaps and failure modes, and turning those findings into actionable insights that directly drive the Agent's iteration at global scale. This includes: