A leading research organization in Cambridge is seeking an intern for a role focused on multimodal algorithmic reasoning. The ideal candidate is a PhD student with experience in machine learning and computer vision, specifically in training multimodal models for vision-and-language tasks. This position offers an opportunity to collaborate with the computer vision team, and candidates are encouraged to bring a strong publication record and experience in mathematical reasoning. Strong programming skills in Python using PyTorch are required.
Qualifications
Strong background in machine learning and computer vision.
Prior experience with training multimodal LLMs for vision-and-language tasks.
Experience in participating and winning mathematical Olympiads.
Responsibilities
Research on multimodal large language models and neural algorithmic reasoning.
Collaborate with researchers in the computer vision team.
Develop algorithms and prepare manuscripts for publications.
Skills
Experience with training large vision-and-language models
Experience with solving mathematical reasoning problems
Programming in Python using PyTorch
Strong track record of publications
Education
Enrolled in a PhD program
Job description
A leading research organization in Cambridge is seeking an intern for a role focused on multimodal algorithmic reasoning. The ideal candidate is a PhD student with experience in machine learning and computer vision, specifically in training multimodal models for vision-and-language tasks. This position offers an opportunity to collaborate with the computer vision team, and candidates are encouraged to bring a strong publication record and experience in mathematical reasoning. Strong programming skills in Python using PyTorch are required.