Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plusai

Santa Clara (CA)

On-site

USD 170,000 - 260,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Catered free lunch
Unlimited snacks and beverages
401(k) plan

Job summary

PlusAI is seeking a talented AI/ML researcher to design, train and deploy Vision-Language-Action models for autonomous trucks. The role focuses on high-level driving decisions, trajectory guidance, and on-vehicle validation across perception, planning and platform teams.

You will own a VLA workstream end-to-end, build robust pipelines, and push state-of-the-art methods into production with distillation and post-training enhancements.

Qualifications

  • MS or PhD in CS/EE/Math/Stats or related field
  • 3+ years implementing and training models in a deep learning framework
  • Hands-on experience training vision-language / vision-language-action models
  • Hands-on experience with model training, evaluation, and deployment in production
  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers
  • Experience with large-scale / distributed model training

Responsibilities

  • Design, train, and evaluate Vision-Language-Action models that generate high-level driving decisions and trajectory guidance in support of Plus's reasoning layer.
  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.
  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.
  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to bring models from research to production

Skills

MS/PhD in CS/EE/Math/Stats
3+ years DL model training
Vision-Language / VLA models
PyTorch TensorFlow JAX
Production deployment experience

Education

MS or PhD in CS/EE/Math/Stats or related field

Tools

PyTorch
TensorFlow
JAX

Job description

PlusAI is a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks. Headquartered in Silicon Valley with operations in the United States and Europe, Plus was named by Fast Company as one of the World’s Most Innovative Companies. Partners including TRATON GROUP’s Scania, MAN, and International brands, Hyundai Motor Company, Iveco Group, Bosch, and DSV are working with Plus to accelerate the deployment of next-generation autonomous trucks. If you’re ready to make a huge impact and drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.


You will join our core AI team at the frontier of autonomous decision-making, building the Vision-Language-Action (VLA) models that form SuperDrive's reasoning layer. You'll train VLA models that generate high-level driving decisions and trajectory guidance for on-board strategic decision-making, and design the knowledge distillation and compression techniques that transition large models onto on-board compute.


Responsibilities


  • Design, train, and evaluate Vision-Language-Action models that generate high-level driving decisions and trajectory guidance in support of Plus's reasoning layer.

  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.

  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.

  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.

  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.

  • Collaborate with perception, planning, and platform teams to bring models from research to production


Required qualifications


  • M.S. minimum, Ph.D. preferred in CS, EE, Mathematics, Statistics, or a related field.

  • 3+ years implementing and training models in a deep learning framework (PyTorch, TensorFlow, or JAX).

  • Direct, hands-on experience training vision-language / vision-language-action models.

  • Hands-on experience with model training, evaluation, and deployment in production.

  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.

  • Experience with large-scale / distributed model training.


Preferred Qualifications


  • Model distillation, quantization, and inference optimization (ONNX/TensorRT, mixed precision, custom kernels).

  • SFT and RL post-training of large multimodal models.

  • Hands-on experience with multi-modal sensor data (camera, LiDAR, radar).

  • Publications at top venues (CVPR, NeurIPS, ICML, ICLR, CoRL, RSS, ICRA).

  • Autonomous driving / ADAS experience.


$170,000 - $260,000 a year


.


Your opportunities joining PlusAI

Work, learn and grow in a highly future-oriented, innovative and dynamic field.


Wide range of opportunities for personal and professional development.


Catered free lunch, unlimited snacks and beverages.


Highly competitive salary and benefits package, including 401(k) plan.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Rival • Santa Clara (CA)

On-site
USD 170,000 - 260,000
Free lunch
Snacks and beverages
401(k) plan
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Socket.dev • Santa Clara (CA)

On-site
USD 150,000 - 210,000
Catered lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plus 2 • Santa Clara (CA)

On-site
USD 150,000 - 190,000
Catered free lunch, unlimited snacks L
Highly competitive salary and benefits
Senior/Staff Machine Learning Engineer, Motion Planning
Senior/Staff Machine Learning Engineer, Motion Planning

Plus 2 • Santa Clara (CA)

On-site
USD 150,000 - 220,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning

PlusAI Inc • Santa Clara (CA)

Hybrid
USD 130,000 - 220,000
Catered lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning

PlusAI • Santa Clara (CA)

On-site
USD 130,000 - 220,000
Catered lunch
Unlimited snacks
401(k) plan
Software Verification & Validation Engineer
Software Verification & Validation Engineer

PlusAI • Santa Clara (CA)

On-site
USD 110,000 - 140,000
401(k) plan
Catered free lunch & snacks
Senior/Staff Machine Learning Engineer, Motion Planning
Senior/Staff Machine Learning Engineer, Motion Planning

PlusAI Inc • Santa Clara (CA)

On-site
USD 130,000 - 220,000
401(k) plan
Free lunch
Senior/Staff Software Engineer (Machine Learning Runtime), Motion Planning
Senior/Staff Software Engineer (Machine Learning Runtime), Motion Planning

PlusAI Inc • Santa Clara (CA)

On-site
USD 130,000 - 220,000
Catered lunch
Unlimited snacks/beverages
401(k) plan
Senior/Staff Software Engineer (Machine Learning Runtime), Motion Planning
Senior/Staff Software Engineer (Machine Learning Runtime), Motion Planning

PlusAI • Santa Clara (CA)

On-site
USD 130,000 - 220,000
Catered free lunch
Unlimited snacks
401(k) plan