Senior Multimodal AI Scientist - Vision & Video Generation

Adobe

San Jose (CA)

On-site

USD 164,000 - 313,300

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

A leading technology company is seeking a qualified candidate for a role focused on enhancing generative AI models. You will design and implement training pipelines, lead development in multimodal areas, and collaborate with various teams to improve quality and efficiency. A Ph.D. in a relevant field is preferred, along with a strong publication record and industry internship experience. The position offers a competitive salary range in California of $216,400–$313,300 annually.

Qualifications

  • Strong publications experience in multimodal generative models.
  • Previous industry-level internship experience is required.
  • Deep understanding of pre-training for multimodal generative models.

Responsibilities

  • Design and implement training pipelines for models.
  • Lead development for pre-training areas for text to image and video.
  • Develop scalable workflows for data quality improvements.

Skills

Pre-training of large-scale multimodal models
Vision-Language Models (VLMs)
Data curation
Distributed training

Education

Ph.D. in Computer Science, Machine Learning, or related field

Tools

Modern diffusion-based architectures (DiT)

Job description

A leading technology company is seeking a qualified candidate for a role focused on enhancing generative AI models. You will design and implement training pipelines, lead development in multimodal areas, and collaborate with various teams to improve quality and efficiency. A Ph.D. in a relevant field is preferred, along with a strong publication record and industry internship experience. The position offers a competitive salary range in California of $216,400–$313,300 annually.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Multimodal AI Researcher - Image/Video Gen & Editing
Multimodal AI Researcher - Image/Video Gen & Editing

Apple Inc. • Santa Clara (CA)

On-site
USD 181,000 - 319,000
Comprehensive medical and dental coverage
Retirement benefits
Employee stock purchase plan
+1
Senior Multimodal AI Scientist - GenAI & Data Pipelines
Senior Multimodal AI Scientist - GenAI & Data Pipelines

Adobe Inc. • San Jose (CA)

On-site
USD 162,000 - 302,000
Senior Multimodal AI Researcher — Flexible, Immersive Media
Senior Multimodal AI Researcher — Flexible, Immersive Media

Dolby Laboratories • Atlanta (GA)

On-site
USD 140,000 - 170,000
Senior AI Researcher - Multimodal Vision, Flexible Work
Senior AI Researcher - Multimodal Vision, Flexible Work

Dolby • Atlanta (GA)

On-site
USD 140,000 - 170,000
Flexible work approach
Health benefits
Bonuses and equity options
Remote Senior AI Researcher: Multimodal Foundation Models
Remote Senior AI Researcher: Multimodal Foundation Models

hum.ai • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Multimodal AI Engineer: Image/Video Generation & Systems
Multimodal AI Engineer: Image/Video Generation & Systems

xAI • Palo Alto (CA)

On-site
USD 180,000 - 440,000
Comprehensive medical, vision, and dental coverage
401(k) retirement plan
Short & long-term disability insurance
+3
Generative CV ML Engineer (Multimodal) — Equity
Generative CV ML Engineer (Multimodal) — Equity

Goliath Partners • San Francisco (CA)

On-site
USD 300,000 - 375,000
Staff Applied Scientist
Staff Applied Scientist

Adobe • San Jose (CA)

On-site
USD 164,000 - 314,000
Senior Multimodal Video AI Researcher
Senior Multimodal Video AI Researcher

Dolby Laboratories • Atlanta (GA)

On-site
USD 125,000 - 163,000
Research Engineer, Multimodal
Research Engineer, Multimodal

character • Redwood City (CA)

On-site
USD 100,000 - 150,000