AI Communication Scientist: Multimodal Robotics & LLMs

VinMotion

Cypress, Northern (CA, KY)

Hybrid

USD 140,000 - 190,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Performance-based bonuses

Job summary

VinMotion USA is seeking an AI Communication Research Scientist to lead research in Generative AI, NLP, and AI Vision for humanoid robotics. This role focuses on developing LLM-driven dialogue systems and multimodal interactions with embodied AI.

Applicants should have a Ph.D. or Master’s in AI/ML/NLP, 3+ years research experience, and strong Python/TensorFlow/PyTorch skills. This full-time position offers a competitive compensation package and opportunities to publish in top conferences.

Qualifications

  • Ph.D. or Master’s degree in AI, NLP, Computer Vision, Machine Learning, Robotics, or related fields.
  • 3+ years of research experience in Generative AI, NLP, speech processing, or AI communication.
  • Strong expertise in transformer-based LLMs, ASR, TTS, and AI Vision.

Responsibilities

  • Conduct research in Generative AI for dialogue generation, speech synthesis, and dynamic response generation.
  • Develop and optimize large language models (GPT, BERT, Whisper, T5, etc.) for real-time communication.
  • Enhance Physical AI by integrating NLP with robot perception, gestures, and emotional intelligence.
  • Implement AI Vision models for lip-reading, visual dialogue grounding, and gesture-based interaction.
  • Work on multimodal AI that fuses text, speech, vision, and physical movement for humanoid robots.
  • Apply Sim2Real techniques to ensure AI models work effectively in real-world robotic environments.
  • Publish research in top-tier conferences (NeurIPS, ACL, ICLR, ICASSP, CVPR).

Skills

Transformers-based LLMs
ASR
TTS
AI Vision

Education

PhD or Master’s degree in AI/NLP/CS/Robotics

Tools

Python
TensorFlow
PyTorch
Hugging Face Transformers

Job description

VinMotion USA is seeking an AI Communication Research Scientist to lead research in Generative AI, NLP, and AI Vision for humanoid robotics. This role focuses on developing LLM-driven dialogue systems and multimodal interactions with embodied AI.

Applicants should have a Ph.D. or Master’s in AI/ML/NLP, 3+ years research experience, and strong Python/TensorFlow/PyTorch skills. This full-time position offers a competitive compensation package and opportunities to publish in top conferences.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Communication Scientist for Humanoid Robots
AI Communication Scientist for Humanoid Robots

VinMotion • Cypress (CA)

On-site
USD 120,000 - 190,000
Competitive compensation
Performance-based bonuses
Work with leading AI researchers
AI Communication Research Scientist - VinMotion US
AI Communication Research Scientist - VinMotion US

VinMotion • Cypress (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Performance-based bonuses
HRI Engineer/Research Scientist
HRI Engineer/Research Scientist

VinMotion • Cypress (CA)

On-site
USD 120,000 - 190,000
Competitive compensation
Performance-based bonuses
Work with leading AI researchers
AI Communication Intern - VinMotion US
AI Communication Intern - VinMotion US

VinMotion • California (MO)

On-site
USD 27,552 - 41,328
Research Scientist
Research Scientist

VinMotion • Cypress (CA), Northern (KY)

Hybrid
USD 120,000 - 170,000
Senior Research Scientist, Humanoid Robotics & Embodied AI
Senior Research Scientist, Humanoid Robotics & Embodied AI

VinMotion • Cypress (CA), Northern (KY)

Hybrid
USD 120,000 - 170,000
Multi-Modal AI Research Scientist – Human Understanding
Multi-Modal AI Research Scientist – Human Understanding

Meta • Pittsburgh, Burlingame (CA)

On-site
USD 150,000 - 190,000
Senior Multimodal AI Scientist — Agentic RL & LLMs
Senior Multimodal AI Scientist — Agentic RL & LLMs

NVIDIA • California (MO)

On-site
USD 152,000 - 242,000
Equity
Benefits
Impact Architect—Generative AI & Multimodal LLMs
Impact Architect—Generative AI & Multimodal LLMs

Cerence AI • United States

Remote
USD 124,000 - 198,000
Annual bonus opportunity
Flexible Time Off
Medical coverage (Health)
+9
Research Engineer - Vision Language Models / Multimodal AI / Computer Vision
Research Engineer - Vision Language Models / Multimodal AI / Computer Vision

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 180,000 - 240,000