Research Scientist, Video Understanding & World Models

Mecka AI

New York (NY)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mecka AI in New York is seeking a Research Scientist specializing in video understanding to lead the development of video representation models. The role involves training large-scale models and fine-tuning video-language models with an emphasis on real-world robotics data.

The ideal candidate will have extensive experience with PyTorch and modern video representation techniques. This position presents a unique opportunity to influence production systems and contribute to cutting-edge research in egocentric embodied video.

Qualifications

  • Deep experience in multi-GPU or distributed training.
  • Strong understanding of multimodal modeling.
  • Experience with egocentric/embodied datasets.

Responsibilities

  • Own model architecture and training strategy.
  • Train and fine-tune video encoders and models.
  • Turn checkpoints into usable artifacts.

Skills

Deep experience training large models in PyTorch
Understanding of modern video representation learning
Ability to run rigorous experiments
Experience with video VLMs
Strong software engineering discipline

Job description

Mecka AI in New York is seeking a Research Scientist specializing in video understanding to lead the development of video representation models. The role involves training large-scale models and fine-tuning video-language models with an emphasis on real-world robotics data.

The ideal candidate will have extensive experience with PyTorch and modern video representation techniques. This position presents a unique opportunity to influence production systems and contribute to cutting-edge research in egocentric embodied video.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist, Video Understanding & World Models
Research Scientist, Video Understanding & World Models

Mecka AI • New York (NY)

On-site
USD 100,000 - 130,000
Research Scientist, Video Understanding & World Models
Research Scientist, Video Understanding & World Models

Kindredventures • New York (NY)

On-site
USD 120,000 - 150,000
Lead Research Scientist, Video Understanding & World Models
Lead Research Scientist, Video Understanding & World Models

Kindredventures • New York (NY)

On-site
USD 120,000 - 150,000
Research Scientist, 3D Perception & Temporal AI
Research Scientist, 3D Perception & Temporal AI

Mecka • New York (NY)

On-site
USD 100,000 - 140,000
Access to proprietary data
Cutting-edge research environment
Opportunity for high impact
Research Scientist: Spatial AI & 3D Reconstruction
Research Scientist: Spatial AI & 3D Reconstruction

Mecka AI • New York (NY)

On-site
USD 120,000 - 150,000
Senior Research Scientist: Multimodal VLM & Video Understanding
Senior Research Scientist: Multimodal VLM & Video Understanding

techire ai • San Francisco (CA)

On-site
USD 210,000 - 260,000
Research Scientist, SLAM & VIO
Research Scientist, SLAM & VIO

Mecka • New York (NY)

On-site
USD 100,000 - 140,000
Access to proprietary data
Cutting-edge research environment
Opportunity for high impact
Machine Learning Engineer (Video Understanding & Segmentation)
Machine Learning Engineer (Video Understanding & Segmentation)

Glint Tech Solutions • Santa Clara (CA)

On-site
USD 140,000 - 210,000
Generative AI Research Engineer - Video World Models
Generative AI Research Engineer - Video World Models

TikTok • San Jose (CA)

On-site
USD 156,000 - 388,000
Research Scientist, Multi-Modal Human Understanding
Research Scientist, Multi-Modal Human Understanding

Meta • Pittsburgh, Burlingame (CA)

On-site
USD 150,000 - 190,000