About Turing:
Turing is one of the world’s fastest-growing AI companies, accelerating the development and deployment of powerful AI systems.
We collaborate with leading AI labs to advance frontier model capabilities in thinking, reasoning, coding, agentic behavior, multimodality, multilinguality, STEM, and frontier knowledge. Our work aims to build real-world AI systems that address mission-critical priorities for companies.
Role Overview:
We are seeking PhD-level Mathematicians to join our frontier AI R&D team. Your focus will be on creating high-quality [problem, solution] data targeting known failure patterns in current top language models. You will work directly with founders to develop mathematical tasks that present deep conceptual challenges, similar to those in FrontierMath or IMO-level problem-solving. Prior experience with competitive math (IMO coaching/problem-setting/winning) is a plus but not mandatory—we are open to developing strong PhD talent in this area.
What You’ll Do:
- Design and solve challenging math problems that reveal weaknesses in large language models.
- Create detailed, step-by-step solutions with clear reasoning and multimodal support (equations, visuals, graphs, simulations).
- Collaborate with founders to align problem types with model evaluation goals, especially in areas where models often fail (e.g., abstraction, multi-step reasoning, symbolic manipulation).
- Help define new evaluation benchmarks inspired by high-school olympiad math, early undergraduate curriculum, and theoretical mathematics.
- Mentor or review work from other math contributors or junior team members.
What We’re Looking For:
- PhD in Mathematics or a related field.
- Strong problem-solving skills, with the ability to design and solve non-standard problems.
- Excellent communication and writing skills, especially in explaining mathematical reasoning step-by-step.
- Familiarity with LaTeX and math visualization tools (e.g., Desmos, GeoGebra, Python/Matplotlib).
- Nice-to-have: Olympiad-level competition experience (contestant, coach, or problem-setter).
- Exposure to AI/LLM research, particularly in evaluation or reasoning tasks.
- Technical setup: Reliable internet, necessary software for computations, visual content creation, and collaboration.
- Ability to work collaboratively with other math enthusiasts.
- Flexible working hours and remote work environment.
Note: Shortlisted candidates may need to complete an assessment during the selection process.