A global technology company at the forefront of artificial intelligence and advanced computing. With a strong commitment to research and innovation, the organisation operates internationally, investing in cutting edge AI technologies and collaborating with leading academic and industry partners to develop next generation intelligent systems.
Our London-based AI research team is expanding and is seeking a Research Scientist – Computer Vision to contribute to pioneering research in multimodal artificial intelligence. Working alongside a highly experienced international team, you will help develop state-of-the-art models spanning computer vision, multimodal learning, and foundation models, with opportunities to translate research into impactful real-world applications.
This is a permanent, full-time position based in Central London.
Key Responsibilities
- Design and develop Vision Transformer (ViT) and multimodal model architectures with enhanced reasoning, efficiency, and scalability
- Advance research in multimodal representation learning, alignment techniques, and long-context modelling
- Investigate scalable training approaches for large multimodal foundation models
- Improve model performance, robustness, and generalisation across diverse tasks
Data & Model Development
- Process and curate large-scale multimodal datasets comprising images, video, audio, and text
- Build robust pipelines for data cleaning, filtering, annotation, and quality assurance
- Maintain reproducible datasets through effective versioning and documentation
- Optimise data sampling strategies and improve dataset quality through iterative evaluation and feedback
Systems & Infrastructure
- Develop distributed training systems for large-scale multimodal models
- Optimise GPU utilisation, resource scheduling, and training efficiency
- Contribute to training and inference frameworks that support scalable model development
- Improve the reliability, performance, and scalability of AI infrastructure
Research Translation
- Apply advanced multimodal AI capabilities to intelligent products and user-facing applications
- Work closely with engineering and product teams to bring research innovations into production
- Contribute to the continuous improvement and deployment of cutting-edge AI technologies
Person specification
Essential
- Degree in Computer Science, Mathematics, Statistics, Artificial Intelligence, or a related technical discipline
- Strong Python programming skills with hands-on experience using PyTorch and modern deep learning frameworks
- Excellent algorithmic thinking, mathematical reasoning, and problem-solving ability
- Strong communication and collaboration skills, with the ability to work effectively across multidisciplinary teams
- A proactive, self-motivated approach and enthusiasm for tackling challenging research problems
Desirable
- Publications at leading AI or computer vision conferences (e.g. CVPR, ICCV, ECCV, NeurIPS, ICML, or ICLR)
- Experience training or fine-tuning large-scale vision, language, or multimodal models
- Contributions to open-source AI projects or research experience within industry or academic laboratories
Please contact Charles Duran for more information.