Research Scientist- Vision-Language-Action (VLA) Models

Robert Bosch Group

Sunnyvale (CA)

On-site

USD 165,000 - 185,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Premium health coverage
401(k) with generous matching
Resources for financial planning and目标

Job summary

Bosch Group in Sunnyvale, California, invites a Research Scientist- Vision-Language-Action (VLA) Models to advance Embodied AI for ADAS/AD, robotics, and industrial automation. You will conduct research and engineer end-to-end perception and planning, and collaborate with global teams to transfer technologies into Bosch platforms.

The role emphasizes publishing results, contributing to conferences, and developing multimodal transformer-based models, with a competitive salary range and strong

Qualifications

  • Ph.D. in Computer Science, Robotics or related discipline, or Master’s with ≥2 years post-graduate R&D experience.
  • Minimum 3 years of R&D experience in AI, CV, robotics or automotive motion planning.
  • Proficiency in Python, C, or Rust for ML/AI workloads.
  • Experience with TensorFlow or PyTorch and RL techniques (PPO, DQN, DDPG).
  • Strong publication record in ML, DL, robotics and CV venues.
  • Strong teamwork and communication skills.

Responsibilities

  • Conduct research and engineering to advance Embodied AI for ADAS/AD and related domains.
  • Advance end-to-end perception and planning with multimodal vision-language-action models.
  • Collaborate with global teams for technology transfer and system integration.
  • Publish findings and/or pursue patents; participate in conferences and workshops.

Skills

Python
C
Rust
Reinforcement learning
Computer Vision
Robotics
Teamwork
Communication

Education

PhD in Computer Science/Robotics or related
Master’s with ≥2 years post-grad R&D

Tools

TensorFlow
PyTorch

Job description

Company Description

The Bosch Research and Technology Center North America with offices in Sunnyvale, California, Pittsburgh, Pennsylvania, and Cambridge, Massachusetts is a part of the global Bosch Group (www.bosch.com), a company with over 70 billion euro revenue, 400,000 employees worldwide, a very diverse product portfolio, and a history spanning over 125 years. The Research and Technology Center North America (RTC-NA) is dedicated to providing technologies and system solutions for various Bosch business fields, primarily in the field of artificial intelligence, energy technologies, internet technologies, circuit design, semiconductors and wireless, as well as advanced MEMS design.


As a part of the global research, our AI research in Silicon Valley focuses on Foundation Models, Big Data Visual Analytics, Explainable AI (XAI), Natural Language Processing, Computer Vision & Mixed Reality, Cloud Robotics, Data Science, AI System Engineering, Time-series Analysis. We develop scalable, intelligent, and trustworthy AIoT solutions for Bosch products and services in application areas such as automated driving, advanced driver assistance systems (ADAS), robotics, smart manufacturing, enterprise AI, health care, smart home and building solutions.


Originating from the AI research in Silicon Valley, our Intelligent Autonomous Systems group is responsible for enabling future autonomous Bosch products by pushing the boundaries of automated driving, advanced driver assistance systems (ADAS), robotics and automation through key innovations that encompass system architecture and AI components. These include methods for motion planning, high level task planning and decision making as well as systems for making these technologies work on real products by building frameworks that take advantage of technologies in the field of reliable distributed computing. We work with internal partners of different Bosch business units to transfer our solutions into future products. We also actively collaborate with leading groups in academia and industry to promote research ideas and publish research findings in internationally renowned conferences and journals such as CVPR, ICRA, IROS, RSS, NeurIPS and CoRL.


Job Description

As a Research Scientist- Vision-Language-Action (VLA) Models, you contribute to research projects at the forefront of the ADAS/AD industry. Key responsibilities include:



  • Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open-world learning, and so on) for related business domains of ADAS/AD, industrial automation, robotics etc.

  • Push the boundaries in (modular) end-to-end perception and planning for ADAS/AD, incorporating advancements in large vision-language-(action) models to aid reasoning capabilities and explainability.

  • Collaborate cross-functionallywith global research and engineering teams to ensure seamless technology transfer and system integration.

  • Implement research results to solve real-world challenges, ensuring high-quality system integration within Bosch’s existing platforms.

  • Stay at the forefront of innovationby actively engaging with academic and industry communities through conferences, workshops, and technical events.

  • Document and disseminate research findings through high-caliber publications and/or patent submissions.


Qualifications

Basic Qualifications


  • Ph.D. in Computer Science, Robotics or a related discipline or Master’s degree with >= 2 years industry experience after graduation.

  • A minimum of 3 years of R&D experience, or an equivalent graduate research background, primarily in AI technologies including Computer Vision and Robotic or Automotive Motion and Behavioral Planning.

  • Proficiency in one or more programming languages commonly used in machine learning (e.g., Python, C , Rust).

  • Strong interpersonal, communication, and teamwork capabilities.

  • Knowledge of major machine learning frameworks like TensorFlow or PyTorch.

  • Hands-on experience in reinforcement learning for behavior or motion planning or other applicable contexts and familiarity with common RL techniques (e.g. PPO, DQN, DDPG).

  • A strong portfolio of publications in premier machine learning, deep learning, robotics and computer vision journals and conferences.


Preferred Qualifications


  • Experience with real-world product development and deployment of autonomous systems.

  • Hands-on experience building and applying multimodal transformer-based sequence-to-sequence models, especially multimodal vision-language-action models.

  • Hands-on experience in computer vision and deep learning, with work in any of the following areas: multimodal transformers, multimodal language models, diffusion models, NeRF, gaussian splatting, object detection / segmentation, 3D scene understanding, sensor calibration, SfM, voxel/BEV grid-based feature representation.


Additional Information

We offer a competitive base salary for this position with a range in US-California of –$165,000 – $185,000 along with an annual corporate bonus, and a long-term incentive bonus designed to reward sustained impact and contribution over time. Within the salary range, the individual pay is determined based on several factors, including, but not limited to, work experience and job knowledge, complexity of the role, job location, etc.


Your well-being matters at Bosch! We offer a benefits package designed to empower you in every area of your life.



  • premium health coverage

  • a 401(k) with generous matching

  • resources for financial planning and goal setting

  • ample paid time off

  • parental leave

  • comprehensive life and disability protection


Learn more about our full benefits offerings by visiting: https://www.myboschbenefits.com/public/welcome.


Equal Opportunity Employer, including disability / veterans.


*Bosch adheres to Federal, State, and Local laws regarding drug-testing. Employment is contingent upon the successful completion of a drug screen and background check. Candidates who have been offered the position must pass both screenings before their start date.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Research Scientist- Vision-Language-Action (VLA) Models
Senior Research Scientist- Vision-Language-Action (VLA) Models

Socket.dev • Sunnyvale (CA)

On-site
USD 185,000 - 215,000
Premium health coverage
401(k) with matching
Paid time off
+2
Research Scientist- Vision-Language-Action (VLA) Models
Research Scientist- Vision-Language-Action (VLA) Models

Bosch USA • Sunnyvale (CA)

On-site
USD 165,000 - 185,000
Health insurance
401(k) matching
Paid time off
+2
Senior Research Scientist- Robotics AI
Senior Research Scientist- Robotics AI

Bosch USA • Sunnyvale (CA)

On-site
USD 185,000 - 215,000
Health coverage
401(k)
Paid time off
+2
Research Scientist VisionLanguageAction VLA Models
Research Scientist VisionLanguageAction VLA Models

Bosch Group • Sunnyvale (CA)

On-site
USD 165,000 - 185,000
Health coverage
401(k) matching
Paid time off
+2
Senior Research Scientist- Robotics AI
Senior Research Scientist- Robotics AI

Socket.dev • Sunnyvale (CA)

On-site
USD 185,000 - 215,000
Research Scientist- Robotics AI
Research Scientist- Robotics AI

Robert Bosch Group • Sunnyvale (CA)

On-site
USD 165,000 - 185,000
Premium health coverage
401(k) with matching
Paid time off
+2
AI Research Scientist- Multimodal Foundational Models
AI Research Scientist- Multimodal Foundational Models

Bosch Group • California (MO)

On-site
USD 165,000 - 195,000
Health coverage
401(k) plan
Paid time off
+2
AI Research Scientist - GenAI
AI Research Scientist - GenAI

Bosch USA • Pittsburgh

On-site
USD 140,000 - 190,000
AI Research Scientist, GenAI
AI Research Scientist, GenAI

Auth21 • Pittsburgh

On-site
USD 120,000 - 170,000
Lead Research Scientist — Vision-Language-Action Models
Lead Research Scientist — Vision-Language-Action Models

Bosch Group • Sunnyvale (CA)

On-site
USD 165,000 - 185,000
Health coverage
401(k) matching
Paid time off
+2