Research Engineer - Generative AI & Multimodal Audio/Video

Google DeepMind

Mountain View (CA)

On-site

USD 174,000 - 252,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Google DeepMind is seeking a research-focused Software Engineer to contribute to large-scale experiments and rapid deployment of AI ideas. You will work across audio, language, and diffusion theory domains, collaborating with interdisciplinary teams to advance Gemini’s audio capabilities and related technologies.

You will bridge theory and code, pursuing independent research initiatives while partnering with universities and publishing results to the broader research community.

Qualifications

  • Bachelor’s degree in Computer Science, Artificial Intelligence, Machine Learning, Computer Vision, Speech Processing, or equivalent practical experience.
  • 3 years of experience in deep learning research and development, including generative AI, audio and video synthesis, diffusion models, and autoregressive generative models.
  • One or more scientific publications in venues like NeurIPS, ICML, ICLR, or equivalent industry contributions.

Responsibilities

  • Be part of team of scientists, engineers, machine learning experts, and more, working together to advance in artificial intelligence.
  • Use our technologies for widespread public benefit and scientific discovery, and collaborate with others on critical issues, ensuring safety and ethics are the highest priority.

Skills

3 years deep learning experience
Generative AI
Diffusion models
Autoregressive models
Publications

Education

Bachelor’s degree in Computer Science or related field

Tools

Python
PyTorch
JAX

Job description

Google DeepMind is seeking a research-focused Software Engineer to contribute to large-scale experiments and rapid deployment of AI ideas. You will work across audio, language, and diffusion theory domains, collaborating with interdisciplinary teams to advance Gemini’s audio capabilities and related technologies.

You will bridge theory and code, pursuing independent research initiatives while partnering with universities and publishing results to the broader research community.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer – AI Audio & Diffusion Systems
Research Engineer – AI Audio & Diffusion Systems

Google • Mountain View (CA)

On-site
USD 174,000 - 253,000
Research Engineer: AI Audio & Diffusion Systems
Research Engineer: AI Audio & Diffusion Systems

Google • United States

On-site
USD 174,000 - 253,000
Research Scientist - Multilingual Audio AI
Research Scientist - Multilingual Audio AI

Google • United States

On-site
USD 181,000 - 245,000
Equity
Bonus target
Research Engineer - Audio & AI Diffusion Innovator
Research Engineer - Audio & AI Diffusion Innovator

Socket.dev • Mountain View (CA)

On-site
USD 174,000 - 252,000
Bonus target
Equity
Benefits
Senior Research Engineer, Multimodal Conversational AI
Senior Research Engineer, Multimodal Conversational AI

Google • New York (NY)

On-site
USD 207,000 - 300,000
Multilingual Audio AI Research Scientist
Multilingual Audio AI Research Scientist

Google • Mountain View (CA)

On-site
USD 174,000 - 252,000
Research Engineer, Multimodal Conversational AI
Research Engineer, Multimodal Conversational AI

Google DeepMind • New York (NY)

Hybrid
USD 207,000 - 300,000
Multimodal Conversational AI Research Engineer
Multimodal Conversational AI Research Engineer

Google • United States

On-site
USD 207,000 - 300,000
Research Engineer, AI Studio & Generative AI (Equity)
Research Engineer, AI Studio & Generative AI (Equity)

Google • New York (NY)

On-site
USD 207,000 - 300,000
Senior Conversational AI Research Engineer - Multimodal NLP
Senior Conversational AI Research Engineer - Multimodal NLP

Google DeepMind • Mountain View (CA)

On-site
USD 207,000 - 300,000
Equity
Bonus target
Benefits