Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Meta seeks a Research Scientist to advance multi-modal AI technologies for human understanding and synthesis. You will develop Vision-Language Models (VLMs) and video foundation models that enable machines to perceive, interpret, and generate rich representations of human behavior and interaction.
Requirements include a Bachelor's degree in a technical field with 2+ years in multi-modal AI, PyTorch experience, and a track record of research contributions at major venues.
Meta is seeking a Research Scientist to advance multi-modal AI technologies for human understanding and synthesis. In this role, you will develop Vision-Language Models (VLMs) and video foundation models that enable machines to perceive, interpret, and generate rich representations of human behavior, expression, and interaction. Your research will span multi-modal reasoning, video understanding, and generative synthesis, enabling more natural and intuitive human-computer interaction at scale.
Currently has, or is in the process of obtaining a Bachelor's degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience. Degree must be completed prior to joining Meta