Research Associate - World Models & Video AI for Robotics

HireArt, Inc.

Los Altos (CA)

On-site

USD 60,000 - 80,000

Part time

31 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Toyota Research Institute (TRI) is hiring a Research Associate to join the Learning from Videos (LFV) team. You will advance world models, video generation, and policy learning, contributing to a shared research codebase and publishing in top venues.

The role emphasizes building state-of-the-art baselines, training large-scale models, and evaluating performance in simulation and on real robotic hardware. The ideal candidate is a current PhD student with strong foundations in computer vision,

Qualifications

  • Current PhD student in a related field.
  • Strong fundamentals in computer vision, video understanding, generative models, 3D reconstruction, or robotics.
  • Experience with video diffusion models, world models, and world-action models.
  • Proactive and self-directed with the ability to work in a research-driven environment.
  • Strong communication and collaboration skills and ownership of problems end-to-end.

Responsibilities

  • Contribute to research on multimodal and multi-view world models, video generation, video policies, and related areas; coordinate with university partnerships, research meetings, and publications.
  • Maintain and evolve the Any4Dv3 video generation codebase, review PRs, stress-test capabilities, and develop new functionality.
  • Benchmark state-of-the-art methods for world-action models (WAMs) within Any4Dv3; deploy in simulation and on real hardware; identify gaps and develop solutions.
  • Produce maintainable, well-documented code and contribute to internal tooling and open-source releases.

Skills

Computer vision
Video understanding
Generative models
3D reconstruction
Robotics

Education

PhD student in related field

Tools

Any4Dv3

Job description

Toyota Research Institute (TRI) is hiring a Research Associate to join the Learning from Videos (LFV) team. You will advance world models, video generation, and policy learning, contributing to a shared research codebase and publishing in top venues.

The role emphasizes building state-of-the-art baselines, training large-scale models, and evaluating performance in simulation and on real robotic hardware. The ideal candidate is a current PhD student with strong foundations in computer vision,

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Research Associate
Research Associate

HireArt, Inc. • Los Altos (CA)

On-site
USD 60,000 - 80,000
Senior Research Engineer, Computer Vision (LFV/WFM)
Senior Research Engineer, Computer Vision (LFV/WFM)

Toyota Research Institute • Los Altos (CA)

On-site
USD 180,000 - 258,750
Medical, dental, and vision insurance
401(k) eligibility
Paid time off benefits
Lead Autonomous Driving Vision Research Engineer
Lead Autonomous Driving Vision Research Engineer

Toyota Research Institute • Los Altos (CA)

On-site
USD 180,000 - 258,750
Medical, dental, and vision insurance
401(k) eligibility
Paid time off benefits
PhD Research Intern — Autonomous Driving & World Models
PhD Research Intern — Autonomous Driving & World Models

Toyota Research Institute • Los Altos (CA)

Hybrid
USD 61,992 - 89,544
Medical insurance
Dental insurance
Vision insurance
+1
Senior Machine Learning Researcher, Large Behavior Models & Diffusion Policy
Senior Machine Learning Researcher, Large Behavior Models & Diffusion Policy

Toyota Research Institute • Los Altos (CA)

On-site
USD 200,000 - 287,500
Medical, dental, vision insurance
401(k) eligibility
Paid time off
+1
Robotics Research Engineer: Vision-Language & World Models
Robotics Research Engineer: Vision-Language & World Models

Relari, inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Postdoctoral Researcher, Human Aware Interaction Learning
Postdoctoral Researcher, Human Aware Interaction Learning

Toyota Research Institute • Cambridge (MA)

On-site
USD 137,000 - 197,000
Medical insurance
Dental insurance
Vision insurance
+4
Research Engineer, Robot Learning
Research Engineer, Robot Learning

Relari, inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000
Senior ML Researcher: End-to-End Driving with LBM & Diffusion
Senior ML Researcher: End-to-End Driving with LBM & Diffusion

Toyota Research Institute • Los Altos (CA)

On-site
USD 200,000 - 287,500
Medical, dental, vision insurance
401(k) eligibility
Paid time off
+1
Human Interactive Driving Intern – World Models
Human Interactive Driving Intern – World Models

Toyota Research Institute • Los Altos (CA)

On-site
USD 61,992 - 89,544
Medical insurance
Dental insurance
Vision insurance
+1