Multimodal AI Researcher: Generative Models, Realtime Vision

Socket.dev

Sunnyvale (CA)

Hybrid

USD 150,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Apple in Sunnyvale is seeking a Multimodal AI Researcher to push the boundaries of foundation models for real-time multimodal data, including video, audio, and text. You will work on interactive models, audio-to-audio modeling, and streaming multimodal systems, driving data requirements, validation strategies, and delivering research that informs product features.

Ideal candidates have a BS with 3+ years of experience, hands-on work with LLMs and VLMs, and strong Python/PyTorch skills.

Qualifications

  • BS and a minimum of 3 years relevant industry experience.
  • Experience building models for multimodal perception systems.
  • Experience working with LLMs and VLMs.
  • Software engineering skills with Python and PyTorch.
  • Curiosity and willingness to learn new things to improve solutions.

Responsibilities

  • Conduct algorithm research and development for multimodal foundational models and agents.
  • Collaborate with data science, ML, and product teams to define requirements and KPIs.
  • Translate research into product features used by millions of users across Apple devices.

Skills

Multimodal AI
Python
PyTorch
LLMs
VLMs
Research experience

Education

BS in related field
MS or PhD in related fields

Job description

Apple in Sunnyvale is seeking a Multimodal AI Researcher to push the boundaries of foundation models for real-time multimodal data, including video, audio, and text. You will work on interactive models, audio-to-audio modeling, and streaming multimodal systems, driving data requirements, validation strategies, and delivering research that informs product features.

Ideal candidates have a BS with 3+ years of experience, hands-on work with LLMs and VLMs, and strong Python/PyTorch skills.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Generative Multimodal AI Researcher
Generative Multimodal AI Researcher

Apple Inc. • Sunnyvale (CA)

Hybrid
USD 150,000 - 278,000
Stock programs
Discretionary bonuses
Relocation
+1
Multimodal AI Researcher
Multimodal AI Researcher

Socket.dev • Sunnyvale (CA)

Hybrid
USD 150,000 - 230,000
Multimodal AI Researcher
Multimodal AI Researcher

Apple Inc. • Sunnyvale (CA)

Hybrid
USD 150,000 - 278,000
Stock programs
Discretionary bonuses
Relocation
+1
Multimodal Image/Video Generation Researcher
Multimodal Image/Video Generation Researcher

Apple Inc. • Seattle (WA)

On-site
USD 184,700 - 324,800
Medical & Dental
Retirement benefits
Employee stock plan
+1
Real-Time Multimodal AI/ML Engineer for Vision
Real-Time Multimodal AI/ML Engineer for Vision

Apple Inc. • Sunnyvale (CA)

Hybrid
USD 150,000 - 278,000
Comprehensive medical and dental cover
Retirement benefits
Discounted products and free services
+3
Multimodal AI Researcher - Agent & Vision-Language Systems
Multimodal AI Researcher - Agent & Vision-Language Systems

Apple Inc. • Santa Clara (CA)

On-site
USD 184,700 - 324,800
Multimodal Machine Learning Researcher
Multimodal Machine Learning Researcher

Socket.dev • Cupertino (CA)

On-site
USD 250,000 - 380,000
Applied AI Scientist - Multimodal Intelligence
Applied AI Scientist - Multimodal Intelligence

Apple • Seattle (WA)

On-site
USD 180,000 - 280,000
Applied AI Scientist - Multimodal Foundation Models
Applied AI Scientist - Multimodal Foundation Models

Apple • Seattle (WA)

On-site
USD 180,000 - 280,000
Multimodal AI Research Engineer (Vision & Language)
Multimodal AI Research Engineer (Vision & Language)

Apple Inc. • Cupertino (CA)

On-site
USD 150,000 - 278,000