Multimodal Vision Researcher – Large Models

microTECH Global Limited

Greater London

Hybrid

GBP 90,000 - 120,000

Full time

14 days+
Application generator

Get a reply from this recruiter — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

microTECH Global Limited in London is pursuing a permanent, full-time role focusing on frontier research in multimodal AI and large language models. You will design ViT-based architectures, advance data pipelines, and develop scalable training systems for real-world deployments.

The role emphasizes strong Python/PyTorch skills, deep learning expertise, and collaboration with cross-functional teams to translate research into production-ready solutions.

Qualifications

  • Bachelor’s degree or higher in Computer Science, Mathematics, Statistics, or related technical field.
  • Strong Python experience with PyTorch and DL frameworks.
  • Excellent algorithm design, mathematical reasoning, and problem solving.
  • Strong communication and collaboration across cross-functional teams.

Responsibilities

  • Develop ViT and multimodal large model architectures with improved reasoning and efficiency.
  • Advance multimodal alignment, representation learning, and long-context modelling.
  • Explore scalable training methods for large multimodal models.
  • Optimize model architectures for generalization and performance.
  • Integrate multimodal capabilities into production and user-facing applications.

Skills

Python
PyTorch
Deep learning
Algorithms
Communication
Cross-functional teamwork
Mathematics

Education

BSc or higher in CS/Math/Stats

Job description

microTECH Global Limited in London is pursuing a permanent, full-time role focusing on frontier research in multimodal AI and large language models. You will design ViT-based architectures, advance data pipelines, and develop scalable training systems for real-world deployments.

The role emphasizes strong Python/PyTorch skills, deep learning expertise, and collaboration with cross-functional teams to translate research into production-ready solutions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Computer Vision Researcher - London (Perm)
Computer Vision Researcher - London (Perm)

microTECH Global Limited • Greater London

Hybrid
GBP 90,000 - 120,000
Multimodal Vision AI Scientist: Research & Impact in London
Multimodal Vision AI Scientist: Research & Impact in London

IC Resources Recruitment • Greater London

On-site
GBP 70,000 - 110,000
Senior Research Engineer - Multimodal & Video Foundation Model
Senior Research Engineer - Multimodal & Video Foundation Model

Tether.io • United Kingdom

On-site
GBP 60,000 - 80,000
Senior AI Research Scientist - Multimodal Models
Senior AI Research Scientist - Multimodal Models

IC Resources • Greater London

On-site
GBP 70,000 - 120,000
Remote AI Research Engineer — LLM & Multimodal
Remote AI Research Engineer — LLM & Multimodal

Tether • United Kingdom

Remote
GBP 111,000 - 141,000
Remote AI Research Engineer - LLM & Multimodal Architect
Remote AI Research Engineer - LLM & Multimodal Architect

Tether.io • United Kingdom

On-site
GBP 90,000 - 150,000
Senior Research Scientist - Multimodal Vision Translation
Senior Research Scientist - Multimodal Vision Translation

The Consensus • Greater London

On-site
GBP 110,000 - 170,000
Senior Multimodal & Video Foundation AI Engineer
Senior Multimodal & Video Foundation AI Engineer

Tether.io • United Kingdom

On-site
GBP 60,000 - 80,000
Elite ML Engineer - Vision & Multimodal AI in Production
Elite ML Engineer - Vision & Multimodal AI in Production

Oho Group • Greater London

On-site
GBP 70,000 - 110,000
Computer Vision Research Scientist
Computer Vision Research Scientist

IC Resources Recruitment • Greater London

On-site
GBP 70,000 - 110,000