Get more replies from employers
Send a job-specific resume in minutes.
Apple seeks a highly skilled researcher to ensure AI features perform across languages and cultures. You will drive multilingual evaluation development from design to implementation, building scalable methods with measurement scientists and ML researchers.
You will publish novel findings while shipping practical evaluation tooling in Python across languages and scripts. You will shape evaluation methodology, test across dialects and cultures, and contribute to scalable pipelines that enable
In this role, you’ll help ensure Apple’s AI features work well across languages and cultures. Your goal is to make our evaluation tooling multilingual from the start so that engineers building AI features can design, test, and ship across the world from day one. It’s a broad applied science role: you’ll shape how Apple evaluates AI wherever the hardest questions are, and you’ll have the opportunity to publish novel work. The scientific challenge is real. How do we ensure we consistently evaluate AI features across different grammar, script, or cultural norms and how do we do this at scale? You’ll bring linguistic judgment to questions like these and, working with measurement scientists and ML researchers, turn it into validated methodology that holds across dozens of languages. This is a hands-on role. You’ll design and implement your own methods in Python, working closely with research and engineering partners, while staying focused on the science of getting evaluation right.