Computer Vision Research Scientist

IC Resources Recruitment

Greater London

On-site

GBP 70,000 - 110,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

IC Resources Recruitment in London is seeking a Computer Vision Research Scientist to join a global technology company at the forefront of AI research and innovation. You will contribute to multimodal AI research, develop state-of-the-art vision transformers and multimodal models, and collaborate with an international team to translate research into real-world applications.

The role offers opportunities to work on large-scale datasets and scalable training systems, with close ties to engineering

Qualifications

  • Bachelor’s degree in Computer Science, Mathematics, Statistics, Artificial Intelligence, or a related technical field.
  • Proficient in Python programming with practical experience in PyTorch and modern deep learning frameworks.
  • Strong algorithmic thinking, mathematical reasoning, and problem-solving skills.
  • Excellent communication and collaboration abilities in multidisciplinary teams.
  • Publications in AI or computer vision conferences are desirable.
  • Experience in training or fine-tuning large-scale vision, language, or multimodal models is a plus.
  • Contributions to open-source AI projects or relevant research experience is advantageous.

Responsibilities

  • Conduct pioneering research in multimodal artificial intelligence, focusing on computer vision.
  • Design and develop Vision Transformer (ViT) and multimodal model architectures for efficiency and scalability.
  • Advance research in multimodal representation learning, alignment techniques, and long-context modeling.
  • Explore scalable training methods for large multimodal foundation models.
  • Enhance model performance, robustness, and generalization across tasks.
  • Process and curate extensive multimodal datasets including images, video, audio, and text.
  • Establish robust data cleaning, filtering, annotation, and quality assurance pipelines.
  • Maintain reproducible datasets through versioning and documentation.
  • Optimize data sampling strategies and improve dataset quality via evaluation and feedback.
  • Develop distributed training systems for large-scale multimodal models.
  • Optimize GPU utilization, resource scheduling, and training efficiency.
  • Contribute to training and inference frameworks that support scalable model development.
  • Enhance reliability, performance, and scalability of AI infrastructure.
  • Apply multimodal AI capabilities to intelligent products and user-facing applications.
  • Collaborate with engineering and product teams to transition research to production.
  • Contribute to ongoing improvement and deployment of cutting-edge AI technologies.

Skills

Python programming
Algorithms
Mathematical reasoning

Education

Bachelor’s degree in Computer Science, Mathematics, Statistics, AI

Tools

PyTorch

Job description

Computer Vision Research Scientist Overview

Company Name: IC Resources Recruitment

Job Role: Computer Vision Research Scientist

Qualifications: Bachelor’s

Category: IT Jobs

Job Type: Full Time

Location: London

An exciting opportunity has arisen for a Computer Vision Research Scientist to join a leading global technology company based in Central London. This organization is at the forefront of artificial intelligence and advanced computing, dedicated to research and innovation while collaborating with top academic and industry partners to develop next-generation intelligent systems.

The successful candidate will be part of an expanding AI research team, contributing to groundbreaking research in multimodal artificial intelligence. You will work alongside a highly skilled international team to develop state-of-the-art models that encompass computer vision, multimodal learning, and foundation models, with the potential to translate research into impactful real-world applications.

Key Responsibilities
  • Conduct pioneering research in multimodal artificial intelligence, focusing on computer vision.
  • Design and develop Vision Transformer (ViT) and multimodal model architectures that enhance reasoning, efficiency, and scalability.
  • Advance research in multimodal representation learning, alignment techniques, and long-context modeling.
  • Explore scalable training methods for large multimodal foundation models.
  • Enhance model performance, robustness, and generalization across various tasks.
  • Process and curate extensive multimodal datasets that include images, video, audio, and text.
  • Establish robust pipelines for data cleaning, filtering, annotation, and quality assurance.
  • Maintain reproducible datasets through effective versioning and documentation.
  • Optimize data sampling strategies and improve dataset quality through iterative evaluation and feedback.
  • Develop distributed training systems for large-scale multimodal models.
  • Optimize GPU utilization, resource scheduling, and training efficiency.
  • Contribute to training and inference frameworks that support scalable model development.
  • Enhance the reliability, performance, and scalability of AI infrastructure.
  • Apply advanced multimodal AI capabilities to intelligent products and user-facing applications.
  • Collaborate closely with engineering and product teams to transition research innovations into production.
  • Contribute to the ongoing improvement and deployment of cutting-edge AI technologies.
Person Specification
  • A degree in Computer Science, Mathematics, Statistics, Artificial Intelligence, or a related technical field.
  • Proficient in Python programming with practical experience in PyTorch and modern deep learning frameworks.
  • Strong algorithmic thinking, mathematical reasoning, and problem-solving skills.
  • Excellent communication and collaboration abilities, capable of working effectively in multidisciplinary teams.
  • A proactive and self-motivated attitude with a passion for addressing challenging research problems.
  • Publications in prominent AI or computer vision conferences such as CVPR, ICCV, ECCV, NeurIPS, ICML, or ICLR are desirable.
  • Experience in training or fine-tuning large-scale vision, language, or multimodal models is a plus.
  • Contributions to open-source AI projects or relevant research experience in industry or academic labs are advantageous.

Degree Requirement: Bachelor’s

Visa Sponsorship May be

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Computer Vision Research Scientist
Computer Vision Research Scientist

IC Resources • Greater London

On-site
GBP 70,000 - 120,000
Computer Vision Engineer
Computer Vision Engineer

microTECH Global LTD • Greater London

On-site
GBP 70,000 - 110,000
Computer Vision Scientist | Multimodal AI / Vision-Language Models / Deep Learning
Computer Vision Scientist | Multimodal AI / Vision-Language Models / Deep Learning

European Tech Recruit • Greater London

On-site
GBP 90,000 - 120,000
Multimodal Vision Research Scientist
Multimodal Vision Research Scientist

IC Resources • Greater London

On-site
GBP 70,000 - 120,000
Research Scientist/Engineer - Multimodal AI & LLM
Research Scientist/Engineer - Multimodal AI & LLM

Adecco • Greater London

On-site
GBP 110,000 - 140,000
Research Scientist/Engineer - Multimodal AI & LLM
Research Scientist/Engineer - Multimodal AI & LLM

Adecco • City Of London

On-site
GBP 85,000 - 120,000
Multimodal Vision AI Scientist: Research & Impact in London
Multimodal Vision AI Scientist: Research & Impact in London

IC Resources Recruitment • Greater London

On-site
GBP 70,000 - 110,000
Computer Vision Research Engineer
Computer Vision Research Engineer

Block MB • Greater London

On-site
GBP 50,000 - 80,000
High autonomy
Direct influence over the research roadmap
Opportunity to publish at top venues
Computer Vision Engineer Lead
Computer Vision Engineer Lead

Connect-AI • England

Hybrid
GBP 127,000 - 150,000
Equity share options
25 days of annual leave plus bank holidays
Comprehensive private medical coverage
+1
Multimodal Vision Engineer — AI Systems & Research
Multimodal Vision Engineer — AI Systems & Research

microTECH Global LTD • Greater London

On-site
GBP 70,000 - 110,000