Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plus 2

Santa Clara (CA)

On-site

USD 150,000 - 190,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Catered free lunch, unlimited snacks L
Highly competitive salary and benefits

Job summary

PlusAI, a Silicon Valley-based Physical AI company, is pushing the frontiers of autonomous trucking with AI-based virtual driver software. Join our core AI team to build Vision-Language-Action models that guide on-board decisions and trajectory planning.

You will own a VLA workstream end-to-end—from data and architecture to large-scale training and in-vehicle validation—while collaborating with perception, planning, and platform teams to deploy models to production.

Qualifications

  • M.S. minimum, Ph.D. preferred in CS, EE, Mathematics, Statistics, or a related field.
  • 3+ years implementing and training models in a deep learning framework.
  • Hands-on experience training vision-language / vision-language-action models.
  • Hands-on experience with model training, evaluation, and deployment in production.
  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.
  • Experience with large-scale / distributed model training.

Responsibilities

  • Design, train, and evaluate Vision-Language-Action models for PlusAI's reasoning layer.
  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.
  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.
  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to bring models from research to production

Skills

Vision-Language-Action (VLA)
PyTorch / TensorFlow / JAX
Distributed training
On-vehicle deployment
SFT / RL post-training
Autonomous driving / ADAS

Education

MS minimum, PhD preferred in CS/EE/Math/Stats

Tools

ONNX/TensorRT
CUDA
Distributed training infra

Job description

PlusAI is a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks. Headquartered in Silicon Valley with operations in the United States and Europe, Plus was named by Fast Company as one of the World’s Most Innovative Companies. Partners including TRATON GROUP’s Scania, MAN, and International brands, Hyundai Motor Company, Iveco Group, Bosch, and DSV are working with Plus to accelerate the deployment of next-generation autonomous trucks. If you’re ready to make a huge impact and drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.

You will join our core AI team at the frontier of autonomous decision-making, building the Vision-Language-Action (VLA) models that form SuperDrive's reasoning layer. You'll train VLA models that generate high-level driving decisions and trajectory guidance for on-board strategic decision-making, and design the knowledge distillation and compression techniques that transition large models onto on-board compute.

  • Design, train, and evaluate Vision-Language-Action models that generate high-level driving decisions and trajectory guidance in support of Plus's reasoning layer.
  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.
  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.
  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to bring models from research to production
  • M.S. minimum, Ph.D. preferred in CS, EE, Mathematics, Statistics, or a related field.
  • 3+ years implementing and training models in a deep learning framework (PyTorch, TensorFlow, or JAX).
  • Direct, hands‑on experience training vision-language / vision-language-action models.
  • Hands‑on experience with model training, evaluation, and deployment in production.
  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.
  • Experience with large-scale / distributed model training.
  • Model distillation, quantization, and inference optimization (ONNX/TensorRT, mixed precision, custom kernels).
  • SFT and RL post-training of large multimodal models.
  • Hands‑on experience with multi‑modal sensor data (camera, LiDAR, radar).
  • Publications at top venues (CVPR, NeurIPS, ICML, ICLR, CoRL, RSS, ICRA).
  • Autonomous driving / ADAS experience.
Your opportunities joining PlusAI

Work, learn and grow in a highly future-oriented, innovative and dynamic field.

Wide range of opportunities for personal and professional development.

  • Catered free lunch, unlimited snacks and beverages.
  • Highly competitive salary and benefits package, including 401(k) plan.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Socket.dev • Santa Clara (CA)

On-site
USD 150,000 - 210,000
Catered lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plusai • Santa Clara (CA)

On-site
USD 170,000 - 260,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Rival • Santa Clara (CA)

On-site
USD 170,000 - 260,000
Free lunch
Snacks and beverages
401(k) plan
Senior Research Engineer, Controls
Senior Research Engineer, Controls

PlusAI Inc • Santa Clara (CA)

On-site
USD 150,000 - 200,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior Research Engineer, Controls
Senior Research Engineer, Controls

PlusAI • Santa Clara (CA)

On-site
USD 150,000 - 200,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Machine Learning Engineer Intern
Machine Learning Engineer Intern

Plus 2 • Santa Clara (CA)

On-site
USD 100,000 - 140,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer, Motion Planning
Senior/Staff Machine Learning Engineer, Motion Planning

PlusAI Inc • Santa Clara (CA)

On-site
USD 130,000 - 220,000
401(k) plan
Free lunch
Software Engineer, Runtime
Software Engineer, Runtime

Plus 2 • Santa Clara (CA)

On-site
USD 120,000 - 180,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning

Plus 2 • Santa Clara (CA)

On-site
USD 120,000 - 180,000
Catered lunch
Unlimited snacks and beverages
401(k) plan
Software Engineer, Runtime
Software Engineer, Runtime

PlusAI • Santa Clara (CA)

On-site
USD 150,000 - 180,000
Catered lunch
Unlimited snacks
401(k) plan
+1