ML Engineer: Vision‑Language Models for Motion

Pear VC

Austin, California (TX, MO)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A cutting-edge AI company in Austin is seeking a Machine Learning Engineer who excels in foundation-model research and production engineering. The role involves training Vision-Language Models to enhance understanding of complex video motion and developing robust APIs for enterprise clients. Ideal candidates will have strong skills in Python and PyTorch, along with research experience in foundational models. Join a dynamic startup environment focused on real-world impact with significant creative autonomy.

Qualifications

  • Strong proficiency in Python and PyTorch.
  • Research experience in foundation models or multi-modal learning.
  • Ability to iterate quickly and autonomously.
  • Experience with video or sensor data training.

Responsibilities

  • Train and evaluate VLMs specialized for motion understanding.
  • Design GPU-accelerated pipelines for multi-modal data.
  • Build frameworks for spatiotemporal reasoning.
  • Develop curation loops that refine datasets.
  • Publish high-impact research while shipping features.

Skills

Python
PyTorch
Multi-modal ML workflows
Autonomous working

Tools

Distributed training systems
GPU optimization

Job description

A cutting-edge AI company in Austin is seeking a Machine Learning Engineer who excels in foundation-model research and production engineering. The role involves training Vision-Language Models to enhance understanding of complex video motion and developing robust APIs for enterprise clients. Ideal candidates will have strong skills in Python and PyTorch, along with research experience in foundational models. Join a dynamic startup environment focused on real-world impact with significant creative autonomy.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Engineer: Vision-Language & Motion for Autonomy
ML Engineer: Vision-Language & Motion for Autonomy

Praxis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 240,000
ML Engineer – Vision-Language for Motion & Autonomy
ML Engineer – Vision-Language for Motion & Autonomy

Nomadic AI • San Francisco (CA)

On-site
USD 170,000 - 250,000
ML Engineer: Motion Planning & Prediction for AVs
ML Engineer: Motion Planning & Prediction for AVs

Avride • Austin (TX)

On-site
USD 100,000 - 130,000
ML Engineer: Vision & AI for Real‑World Systems
ML Engineer: Vision & AI for Real‑World Systems

Kitware • Town of Clifton Park (NY)

On-site
USD 85,000 - 125,000
Flexible working hours
401(k)
Health insurance
+3
ML Engineer: Vision & Generative AI Systems
ML Engineer: Vision & Generative AI Systems

Apple Inc. • Los Angeles (CA)

On-site
USD 171,000 - 303,000
Comprehensive medical and dental coverage
Retirement benefits
Employee stock programs
+1
AI Engineer - Robotics
AI Engineer - Robotics

Jobzhr • San Francisco (CA)

On-site
USD 150,000 - 190,000
ML Engineer: On-Device AI & Vision Systems
ML Engineer: On-Device AI & Vision Systems

Apple Inc. • Seattle (WA)

On-site
USD 139,000 - 259,000
ML Engineer — Production-Grade AI & LLM/VLM Systems
ML Engineer — Production-Grade AI & LLM/VLM Systems

Nace.AI • Palo Alto (CA)

On-site
USD 120,000 - 160,000
ML Engineer II — Computer Vision & Generative AI
ML Engineer II — Computer Vision & Generative AI

SMS Superior Maintenance Solutions • New York (NY)

Hybrid
USD 120,000 - 150,000
ML Engineer – Real-Time Audio AI & Production Systems
ML Engineer – Real-Time Audio AI & Production Systems

Catalyst Labs • Austin (TX)

On-site
USD 90,000 - 130,000