A complete application in a minute — tailored resume and cover letter, ready to send.
MBZUAI invites applications for a Research Scientist role in the Vision Language Model team to advance state-of-the-art multimodal foundation models. You will research architectures, data recipes, and benchmarks for large-scale VLM systems, integrating visual understanding, language reasoning, and agentive capabilities.
You will collaborate with world-class researchers and engineers, publish results, mentor junior researchers, and help scale training pipelines with reinforcement learning
MBZUAI invites applications for a Research Scientist role in the Vision Language Model team to advance state-of-the-art multimodal foundation models. You will research architectures, data recipes, and benchmarks for large-scale VLM systems, integrating visual understanding, language reasoning, and agentive capabilities.
You will collaborate with world-class researchers and engineers, publish results, mentor junior researchers, and help scale training pipelines with reinforcement learning