A complete application in a minute — tailored resume and cover letter, ready to send.
Storm3 in San Francisco is seeking researchers with expertise in VLMs, vision-language, and multimodal understanding to advance the capabilities of our LLMs.
PhD in CS/ML is preferred, with 2+ years of industry experience in VLMs and a strong publication record. You will help build the multimodal data infrastructure, curate datasets, and publish breakthroughs with the team.
Come join one of the only research institutions globally with resources to compete with top AI companies => 10s of 1000s of GPUs and Tier1 talent.
This lab is a playground for state-of-the-art research in LLMs, Agents and World Models.
Hiring for those experienced in VLMs, vision-language, image/video understanding & reasoning to contribute to the multimodal capabilities of their LLMs.