Senior Multimodal AI Scientist - Vision & Video Generation
Adobe
San Jose (CA)
On-site
USD 164,000 - 313,300
Full time
14 days+
Application generator
Stand out for this role — generate a tailored resume and cover letter in about a minute.
Get past ATS filters
Job summary
A leading technology company is seeking a qualified candidate for a role focused on enhancing generative AI models. You will design and implement training pipelines, lead development in multimodal areas, and collaborate with various teams to improve quality and efficiency. A Ph.D. in a relevant field is preferred, along with a strong publication record and industry internship experience. The position offers a competitive salary range in California of $216,400–$313,300 annually.
Qualifications
Strong publications experience in multimodal generative models.
Previous industry-level internship experience is required.
Deep understanding of pre-training for multimodal generative models.
Responsibilities
Design and implement training pipelines for models.
Lead development for pre-training areas for text to image and video.
Develop scalable workflows for data quality improvements.
Skills
Pre-training of large-scale multimodal models
Vision-Language Models (VLMs)
Data curation
Distributed training
Education
Ph.D. in Computer Science, Machine Learning, or related field
Tools
Modern diffusion-based architectures (DiT)
Job description
A leading technology company is seeking a qualified candidate for a role focused on enhancing generative AI models. You will design and implement training pipelines, lead development in multimodal areas, and collaborate with various teams to improve quality and efficiency. A Ph.D. in a relevant field is preferred, along with a strong publication record and industry internship experience. The position offers a competitive salary range in California of $216,400–$313,300 annually.