GenAI Multimodal Scientist — Prompt Engineer

LTM

San Jose (CA)

On-site

USD 150,000 - 230,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Medical plan
Disability coverage
401(k) plan with company match
Life insurance
Paid time off
Parental leave

Job summary

LTIMindtree is seeking Applied Scientists in the United States (San Jose) to design and optimize Gen AI driven multimodal pipelines for detecting contextual moments in live video streams. The role emphasizes prompt engineering, multimodal inference, and model optimisation using Claude Nova Bedrock to maximise detection accuracy.

You will work on building pipelines for multimodal inference, video transcripts, and audio, with a focus on structured JSON metadata and integration into downstream

Qualifications

  • Experience with prompt engineering for foundation models.
  • Proficiency in multimodal AI systems (vision and NLP).
  • Strong Python programming and data-pipeline skills.
  • Experience with LLM inference and optimization.
  • Familiarity with AWS cloud services (Bedrock, Lambda, S3) and analytics pipelines.

Responsibilities

  • Design and optimize prompt engineering strategies for foundation models.
  • Build pipelines for multimodal inference (video, transcript, audio).
  • Detect and classify contextual moments and events.
  • Develop reusable frameworks and moment-detection templates.
  • Generate structured metadata in JSON for downstream systems.

Skills

Prompt engineering
Multimodal AI
LLM inference
Python programming
Data pipelines
Video analytics

Tools

AWS Lambda
Amazon S3
Bedrock

Job description

LTIMindtree is seeking Applied Scientists in the United States (San Jose) to design and optimize Gen AI driven multimodal pipelines for detecting contextual moments in live video streams. The role emphasizes prompt engineering, multimodal inference, and model optimisation using Claude Nova Bedrock to maximise detection accuracy.

You will work on building pipelines for multimodal inference, video transcripts, and audio, with a focus on structured JSON metadata and integration into downstream

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

GenAI Multimodal Scientist: Prompting & Video Analytics
GenAI Multimodal Scientist: Prompting & Video Analytics

LTM • Edison (NJ)

On-site
USD 120,000 - 190,000
Comprehensive Medical Plan
401(k) Plan with Company match
Paid Holidays
Data Scientist
Data Scientist

LTM • Edison (NJ)

On-site
USD 120,000 - 190,000
Comprehensive Medical Plan
401(k) Plan with Company match
Paid Holidays
Senior AI UX Designer: Multimodal & Prompt Prototyping
Senior AI UX Designer: Multimodal & Prompt Prototyping

Google DeepMind • Mountain View (CA)

On-site
USD 160,000 - 231,000
Health insurance
Dental insurance
Vision insurance
+8
Multimodal AI Engineer: Vision, Audio & Text
Multimodal AI Engineer: Vision, Audio & Text

AI Breaking Wire • Mountain View (CA), Northern (KY)

Hybrid
USD 210,000 - 310,000
Equity program
Healthcare
On-site gym
+3
Multimodal Intelligence Systems Engineer
Multimodal Intelligence Systems Engineer

MaxIT Consulting - Max Corporate Group • San Francisco (CA)

On-site
USD 120,000 - 180,000
Senior Specialist - Data Sciences
Senior Specialist - Data Sciences

LTM • San Jose (CA)

On-site
USD 150,000 - 230,000
Medical plan
Disability coverage
401(k) plan with company match
+3
Multimodal AI Research Engineer for Autonomous Robotics
Multimodal AI Research Engineer for Autonomous Robotics

Acceler8 Talent • Mountain View (CA)

Hybrid
USD 180,000 - 240,000
Architect ML/GenAI IRC296973
Architect ML/GenAI IRC296973

GlobalLogic • Town of Poland (NY)

On-site
USD 120,000 - 160,000
Multimodal AI Research Scientist
Multimodal AI Research Scientist

Thinking Machines Lab Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3
Senior AI Engineer: Multimodal & LLM Orchestration (Remote)
Senior AI Engineer: Multimodal & LLM Orchestration (Remote)

Embedded Shishya • United States

Remote
USD 120,000 - 190,000
Remote work within LatAm
Wellness stipend
AI Voucher
+1