Senior Algorithm Engineer (Multimodal Generation)

PERSOL SINGAPORE PTE. LTD.

Singapore

On-site

SGD 90,000 - 150,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PERSOL SINGAPORE PTE. LTD. is seeking a researcher/developer to advance multimodal AI capabilities, including image and video generation and editing, and to push unified generation‑and‑understanding models.

The role requires a Master’s in CS/AI, 1–5 years of relevant work, strong Python/C++, and experience with PyTorch or TensorFlow. You will collaborate with product, system, and engineering teams to deploy innovations into real-world solutions.

Qualifications

  • Master’s degree or above in Computer Science, Artificial Intelligence, or related engineering fields.
  • 1–5 years of relevant research or development experience in image/video generation or multimodal models.
  • Strong Python/C++ skills and experience with PyTorch or TensorFlow.
  • First-author publications in top conferences a strong plus (CVPR, ICCV, ECCV, NeurIPS, ACMMM, ACL, ICLR).
  • Ability to independently conduct model design, training, and experimental analysis.

Responsibilities

  • R&D on large multimodal models (LMMs) and AIGC, focusing on image/video generation and editing and related capabilities.
  • Evaluate state-of-the-art video generation and multimodal models; incubate new capabilities across devices, cloud, and automotive.
  • Collaborate with product, system, and engineering teams to translate innovations into deployable product solutions.

Skills

Python
C++
PyTorch
TensorFlow
Deep learning
Research

Education

Master’s degree in CS/AI

Tools

NumPy

Job description

Job Responsibilities:
  • Conduct research and development on large multimodal models (LMMs) and AIGC generation algorithms, with a focus on image and video generation and editing, continuously advancing key algorithmic capabilities including but not limited to multimodal controllable generation, video continuation, image and video editing, and unified generation-and-understanding models;
  • Explore and evaluate state-of-the-art video generation and unified generation–understanding models, closely tracking advances from both academia and industry in image, video, and multimodal large models, and incubate new capabilities, features, and application directions across device, cloud, and autonomous driving scenarios;
  • Work closely with product, system, and engineering teams to continuously translate cutting‑edge algorithmic innovations into deployable and deliverable product solutions, ensuring stable application of related technologies in real‑world business scenarios.
Job Requirements:
  • Master’s degree or above in Computer Science,Artificial Intelligence, or related engineering fields
  • 1–5 years of relevant research or development experience, with solid foundations in one or more of the following areas: image/video generation and editing, multimodal large models, unified generation-and-understanding models, reinforcement learning, etc.; First-author publications in top‑tier conferences or journals are a strong plus, such as CVPR, ICCV, ECCV, NeurIPS, ACMMM, ACL, ICLR, etc.;
  • Strong algorithm and deep learning model development skills, proficiency in Python / C++, and hands‑on experience with mainstream AI frameworks such as PyTorch or TensorFlow, with the ability to independently conduct model design, training, and experimental analysis;
  • Experience in image or video generation and editing algorithms is a plus; practical experience with large‑scale visual data annotation, cleaning, and analysis is an added advantage;
  • Broad technical vision with the ability to continuously follow and understand the latest advances in multimodal and generative AI from both academia and industry;
  • Strong independent thinking skills and a collaborative team‑oriented mindset.

We regret to inform that only shortlisted candidates will be notified.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Multimodal AI Engineer — Video & Generative Models
Senior Multimodal AI Engineer — Video & Generative Models

PERSOL SINGAPORE PTE. LTD. • Singapore

On-site
SGD 90,000 - 150,000
AIGC Video Generation Algorithm (Leader)
AIGC Video Generation Algorithm (Leader)

Shopee • Singapore

On-site
SGD 180,000 - 260,000
Machine Learning Researcher, Image/ Video Generation
Machine Learning Researcher, Image/ Video Generation

Apple • Singapore

On-site
SGD 70,000 - 100,000
Machine Learning Researcher, Image/ Video Generation
Machine Learning Researcher, Image/ Video Generation

Lex • Singapore

On-site
SGD 120,000 - 180,000
AI Algo Eng – LLM/VLM (Mandarin Required)
AI Algo Eng – LLM/VLM (Mandarin Required)

Applied Intelligence Consulting (Singapore) • Singapore

On-site
SGD 120,000 - 180,000
Intelligent Video Processing Algorithm Engineer
Intelligent Video Processing Algorithm Engineer

PERSOL SINGAPORE PTE. LTD. • Singapore

On-site
SGD 70,000 - 110,000
Image Signal Processing Algorithm Engineer
Image Signal Processing Algorithm Engineer

PERSOL SINGAPORE PTE. LTD. • Singapore

On-site
SGD 120,000 - 180,000
AI ENGINEER
AI ENGINEER

BEYONDSOFT CONSULTING (SINGAPORE) PTE. LTD. • Singapore

On-site
SGD 100,000 - 140,000
Machine Learning Researcher, Image/ Video Generation
Machine Learning Researcher, Image/ Video Generation

Apple Inc. • Singapore

On-site
SGD 110,000 - 170,000
Multimodal Vision & Image/Video Gen AI Researcher
Multimodal Vision & Image/Video Gen AI Researcher

Apple Inc. • Singapore

On-site
SGD 110,000 - 170,000