Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Socket.dev

Santa Clara (CA)

On-site

USD 150,000 - 210,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Catered lunch
Unlimited snacks and beverages
401(k) plan

Job summary

PlusAI, a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks, invites talented researchers to join its core AI team in Silicon Valley.

You will train Vision-Language-Action models that generate high-level driving decisions and trajectory guidance for on-board decision-making, collaborating across perception, planning, and platform teams to bring models from research to production.

Qualifications

  • MS minimum, PhD preferred in CS, EE, Mathematics, Statistics, or related field.
  • 3+ years implementing and training models in a deep learning framework (PyTorch, TensorFlow, or JAX).
  • Direct, hands-on experience training vision-language / vision-language-action models.
  • Hands-on experience training, evaluation, and deployment in production.
  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.
  • Experience with large-scale / distributed model training.

Responsibilities

  • Design, train, and evaluate Vision-Language-Action models that generate high-level driving decisions and trajectory guidance in support of Plus's reasoning layer.
  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.
  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.
  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to bring models from research to production

Skills

Vision-Language
DL Frameworks
Distributed Training

Education

MS or PhD in CS/EE/Math

Tools

PyTorch
TensorFlow
JAX
ONNX/TensorRT

Job description

PlusAI is a Physical AI company pioneering AI-based virtual driver software for factory-built autonomous trucks. Headquartered in Silicon Valley with operations in the United States and Europe, Plus was named by Fast Company as one of the World’s Most Innovative Companies. Partners including TRATON GROUP’s Scania, MAN, and International brands, Hyundai Motor Company, Iveco Group, Bosch, and DSV are working with Plus to accelerate the deployment of next-generation autonomous trucks. If you’re ready to make a huge impact and drive the future of autonomy, Plus is looking for talented individuals to join its fast-growing teams.

You will join our core AI team at the frontier of autonomous decision-making, building the Vision‑Language‑Action (VLA) models that form SuperDrive's reasoning layer. You'll train VLA models that generate high-level driving decisions and trajectory guidance for on-board strategic decision-making, and design the knowledge distillation and compression techniques that transition large models onto on-board compute.

Responsibilities
  • Design, train, and evaluate Vision-Language-Action models that generate high-level driving decisions and trajectory guidance in support of Plus's reasoning layer.
  • Own a VLA workstream end to end — data, architecture, large-scale training, and on-vehicle validation.
  • Build training and evaluation pipelines and rigorous metrics for VLA performance in driving contexts.
  • Develop distillation and compression recipes to deploy large reasoning models on on-board compute.
  • Apply SFT and RL post-training to improve reasoning, robustness, and long-tail behavior.
  • Collaborate with perception, planning, and platform teams to bring models from research to production
Required qualifications
  • M.S. minimum, Ph.D. preferred in CS, EE, Mathematics, Statistics, or a related field.
  • 3+ years implementing and training models in a deep learning framework (PyTorch, TensorFlow, or JAX).
  • Direct, hands‑on experience training vision-language / vision-language-action models.
  • Hands‑on experience with model training, evaluation, and deployment in production.
  • Thorough understanding of state-of-the-art vision-language / VLA models, diffusion, flow matching, and transformers.
  • Experience with large-scale / distributed model training.
Preferred Qualifications
  • Model distillation, quantization, and inference optimization (ONNX/TensorRT, mixed precision, custom kernels).
  • SFT and RL post-training of large multimodal models.
  • Hands‑on experience with multi-modal sensor data (camera, LiDAR, radar).
  • Publications at top venues (CVPR, NeurIPS, ICML, ICLR, CoRL, RSS, ICRA).
  • Autonomous driving / ADAS experience.
Your opportunities joining PlusAI

Work, learn and grow in a highly future-oriented, innovative and dynamic field.

Wide range of opportunities for personal and professional development.

  • Catered free lunch, unlimited snacks and beverages.
  • Highly competitive salary and benefits package, including 401(k) plan.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plus 2 • Santa Clara (CA)

On-site
USD 150,000 - 190,000
Catered free lunch, unlimited snacks L
Highly competitive salary and benefits
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Rival • Santa Clara (CA)

On-site
USD 170,000 - 260,000
Free lunch
Snacks and beverages
401(k) plan
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)
Senior/Staff Research Engineer — Vision-Language-Action Models (Autonomous Driving)

Plusai • Santa Clara (CA)

On-site
USD 170,000 - 260,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior Research Engineer, Controls
Senior Research Engineer, Controls

PlusAI • Santa Clara (CA)

On-site
USD 150,000 - 200,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior Research Engineer, Controls
Senior Research Engineer, Controls

PlusAI Inc • Santa Clara (CA)

On-site
USD 150,000 - 200,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer, Motion Planning
Senior/Staff Machine Learning Engineer, Motion Planning

PlusAI Inc • Santa Clara (CA)

On-site
USD 130,000 - 220,000
401(k) plan
Free lunch
Machine Learning Engineer Intern
Machine Learning Engineer Intern

Plus 2 • Santa Clara (CA)

On-site
USD 100,000 - 140,000
Catered free lunch
Unlimited snacks and beverages
401(k) plan
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning
Senior/Staff Machine Learning Engineer (Reinforcement Learning), Motion Planning

Plus 2 • Santa Clara (CA)

On-site
USD 120,000 - 180,000
Catered lunch
Unlimited snacks and beverages
401(k) plan
Senior Simulation Software Engineer
Senior Simulation Software Engineer

Plusai • Santa Clara (CA)

On-site
USD 130,000 - 200,000
Free lunch
Snacks & beverages
401(k) plan
Software Engineer, Runtime
Software Engineer, Runtime

PlusAI • Santa Clara (CA)

On-site
USD 150,000 - 180,000
Catered lunch
Unlimited snacks
401(k) plan
+1