COMPUTER VISION ENGINEER (LLM & AI Integration)

Duncan & Ross Consulting

Abu Dhabi

On-site

AED 331,000 - 477,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Duncan & Ross Consulting seeks a Computer Vision Engineer to design, build, and deploy multimodal vision solutions that interpret and describe visual data for integrated AI applications.

You will collaborate with data scientists to fine-tune vision-language models, develop production pipelines, and deploy code using PyTorch, TensorFlow, and OpenCV while upholding ethical AI practices and data privacy.

Qualifications

  • 3–7 years of experience in computer vision, deep learning, or multimodal AI.
  • Strong proficiency in Python and frameworks such as PyTorch, TensorFlow, Keras, and OpenCV.
  • Experience integrating LLMs (GPT, Claude, Gemini, or open-source models) with vision systems.
  • Solid understanding of transformer architectures, CNNs, diffusion models, and attention mechanisms.
  • Familiarity with multimodal datasets (COCO, Visual Genome) and evaluation metrics for vision tasks.
  • Experience with cloud-based AI tools (Azure AI, AWS Sagemaker, Google Vertex AI).
  • Ability to write clean, scalable, production-grade code.
  • Strong analytical, problem-solving, and communication skills.

Responsibilities

  • Develop and implement computer vision models for image classification, object detection, segmentation, facial recognition, and visual understanding.
  • Integrate vision models with LLMs to build systems that interpret and describe visual content.
  • Design AI pipelines that combine text, images, and video data for multimodal learning and reasoning.
  • Utilize deep learning frameworks (TensorFlow, PyTorch, OpenCV) to prototype and deploy models.
  • Collaborate with data scientists and AI researchers to fine-tune vision-language models for specific tasks such as visual QA, captioning, or scene analysis.
  • Implement data preprocessing, augmentation, and annotation pipelines for large-scale image datasets.
  • Conduct performance benchmarking, optimization, and deployment of models in production environments.
  • Research and experiment with emerging techniques in Generative AI, multimodal transformers, and neural architecture optimization.
  • Develop APIs and tools for internal teams to utilize vision + LLM capabilities.
  • Ensure compliance with ethical AI practices, including bias mitigation and data privacy.

Skills

Python
PyTorch
TensorFlow
OpenCV
LLM integration
Multimodal AI
Transformer models
Cloud platforms

Education

Bachelor's or Master’s in CS/AI
PhD preferred

Tools

PyTorch
TensorFlow
Keras
OpenCV
ONNX

Job description

COMPUTER VISION ENGINEER (LLM & AI Integration)
About the job COMPUTER VISION ENGINEER (LLM & AI Integration)

JOB SUMMARY:

We are seeking an experienced Computer Vision Engineer with a strong background in AI and Large Language Models (LLMs). The ideal candidate will design, build, and deploy computer vision solutions that integrate with generative AI and LLM frameworks to interpret, analyze, and describe visual data. This role bridges the gap between image understanding and natural language processing, enabling intelligent visual-language applications.

KEY RESPONSIBILITIES:

  • Develop and implement computer vision models for image classification, object detection, segmentation, facial recognition, and visual understanding.
  • Integrate vision models with LLMs (e.g., GPT, LLaVA, CLIP, or multimodal models) to build systems that interpret and describe visual content.
  • Design AI pipelines that combine text, images, and video data for multimodal learning and reasoning.
  • Utilize deep learning frameworks (TensorFlow, PyTorch, OpenCV) to prototype and deploy models.
  • Collaborate with data scientists and AI researchers to fine-tune vision-language models for specific tasks such as visual QA, captioning, or scene analysis.
  • Implement data preprocessing, augmentation, and annotation pipelines for large-scale image datasets.
  • Conduct performance benchmarking, optimization, and deployment of models in production environments.
  • Research and experiment with emerging techniques in Generative AI, multimodal transformers, and neural architecture optimization.
  • Develop APIs and tools for internal teams to utilize vision + LLM capabilities.
  • Ensure compliance with ethical AI practices, including bias mitigation and data privacy.

QUALIFICATIONS:

  • Bachelors or Masters degree in Computer Science, AI, Computer Vision, or related field (PhD preferred).
  • 3-7 years of experience in computer vision, deep learning, or multimodal AI.
  • Strong proficiency in Python and frameworks such as PyTorch, TensorFlow, Keras, and OpenCV.
  • Experience integrating LLMs (GPT, Claude, Gemini, or open-source models) with vision systems.
  • Solid understanding of transformer architectures, CNNs, diffusion models, and attention mechanisms.
  • Familiarity with multimodal datasets (COCO, Visual Genome, etc.) and evaluation metrics for vision tasks.
  • Experience with cloud-based AI tools (Azure AI, AWS Sagemaker, Google Vertex AI, etc.).
  • Ability to write clean, scalable, and production-grade code.
  • Strong analytical, problem-solving, and communication skills.

PREFERRED QUALIFICATIONS:

  • Experience with multimodal LLM frameworks such as CLIP, BLIP, LLaVA, or Kosmos-2.
  • Background in natural language processing and prompt engineering.
  • Hands-on experience with edge deployment (NVIDIA Jetson, OpenVINO, ONNX).
  • Knowledge of reinforcement learning, generative models, or 3D vision.
  • Publications or open-source contributions in AI research are a plus.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Specialist- Native Arabic Speakers
AI Specialist- Native Arabic Speakers

Confidential Company • Abu Dhabi

On-site
AED 320,000 - 520,000
Multimodal Vision AI Engineer (LLM Integration)
Multimodal Vision AI Engineer (LLM Integration)

Duncan & Ross Consulting • Abu Dhabi

On-site
AED 331,000 - 477,000
Specialist Artificial Intelligence
Specialist Artificial Intelligence

General Civil Aviation Authority • Abu Dhabi

On-site
AED 220,000 - 320,000
AI Engineer - Convo AI
AI Engineer - Convo AI

Dicetek LLC • Abu Dhabi

On-site
AED 180,000 - 250,000
Data Scientist- Computer Vision (Arabic Speakers)
Data Scientist- Computer Vision (Arabic Speakers)

Addenda • United Arab Emirates

On-site
AED 391,000 - 670,000
Lead AI Scientist / Head of AI Solutions
Lead AI Scientist / Head of AI Solutions

EstateSight AI • Abu Dhabi

On-site
AED 450,000 - 900,000
Opportunity to lead AI innovation with societal impact
Research-driven environment
Leadership exposure across teams
Machine Learning Engineer – Generative AI (LLMs / RAG / Agentic AI)
Machine Learning Engineer – Generative AI (LLMs / RAG / Agentic AI)

Stellar Technologies • Abu Dhabi

On-site
AED 260,000 - 480,000
Lead AI Scientist / Head of AI Solutions
Lead AI Scientist / Head of AI Solutions

Recenso • Abu Dhabi

On-site
Opportunity to lead AI innovation
Research-driven environment
Leadership exposure across teams
AI Engineer (Vietnamese speaker)
AI Engineer (Vietnamese speaker)

AGAPI Technologies • United Arab Emirates

On-site
AED 180,000 - 280,000
Competitive compensation
Visa processing and UAE benefits
Generous holidays
+3
Software Development Engineer II
Software Development Engineer II

NextGenEnergyJobs • Sharjah

On-site
AED 180,000 - 300,000