Get more replies from employers
Send a job-specific resume in minutes.
European Tech Recruit is partnering with a London-based leader in AI research to hire a Research Scientist in Computer Vision and multimodal AI. The role focuses on next‑generation vision-language models, large-scale multimodal systems, and production-ready AI applications.
You will explore ViT architectures, cross-modal understanding, and scalable training across distributed GPUs, while curating huge multimodal datasets and building robust data pipelines.
Computer Vision Scientist | Multimodal AI / Vision-Language Models / Deep Learning
We are currently partnered with a leading global technology company with a major research and development presence in London, focused on advancing the frontiers of artificial intelligence, computer vision, multimodal understanding, and embodied intelligence.
As part of their continued investment in fundamental and applied AI research, they are looking to hire a Research Scientist specialising in Computer Vision and multimodal AI to contribute to next-generation vision-language models, large-scale multimodal learning systems, and production-ready AI applications.
Research Scientist / Computer Vision Scientist / AI Research Scientist / Machine Learning Researcher / Computer Vision Researcher / Multimodal AI / Multimodal Machine Learning / Vision-Language Models / VLM / Multimodal Large Language Models / MLLM / Large Language Models / LLM / Vision Transformer / ViT / Deep Learning / PyTorch / Python / Foundation Models / Generative AI / Representation Learning / Multimodal Alignment / Long Context / Computer Vision / Image Understanding / Video Understanding / AI Research / Distributed Training / GPU Computing / Model Pre-Training / Fine-Tuning / Model Serving / AI Infrastructure / Data Curation / Dataset Development / CVPR / ICCV / ECCV / NeurIPS / ICML / ICLR / London / UK AI Research
By applying to this role, you understand that we may collect your personal data and process it in line with our privacy policy.