On-Device Speech & Multimodal Scientist

HushOne, Inc.

Kirkland (WA)

On-site

USD 120,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Stock options
Annual bonus
Health, dental, vision insurance
401(k) with company match
AI tokens
Gym membership
Referral program

Job summary

HushOne, Inc. is seeking a researcher or engineer to develop and adapt models for speech, documents, images and grounded interaction on-device. You will evaluate diverse accents, environments and accessibility needs with participant consent, under real device constraints.

In the first 90 days, deliver a bounded multimodal capability with clear limitations and a path to production. Collaboration across product and hardware teams is essential, with emphasis on practical evaluation and

Qualifications

  • Research or advanced engineering in speech, vision, or multimodal learning.
  • You can show how you distinguish apparent fluency from correct understanding.
  • You work within real device constraints rather than assuming a server.
  • You evaluate on realistic input, including accents, noise, and bad lighting.

Responsibilities

  • Develop or adapt models for speech, documents, images and grounded interaction.
  • Evaluate diverse accents, environments and accessibility needs with consent.
  • Measure local performance and make capture, retention and activation behavior understandable.
  • Design evaluation for a voice-controlled task in a shared room.

Skills

On-device multimodal research
Speech/vision learning
Evaluation under real device limits
Distinguish fluency from understanding

Job description

HushOne, Inc. is seeking a researcher or engineer to develop and adapt models for speech, documents, images and grounded interaction on-device. You will evaluate diverse accents, environments and accessibility needs with participant consent, under real device constraints.

In the first 90 days, deliver a bounded multimodal capability with clear limitations and a path to production. Collaboration across product and hardware teams is essential, with emphasis on practical evaluation and

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote‑Friendly Builders & Researchers — Open Application
Remote‑Friendly Builders & Researchers — Open Application

HushOne, Inc. • Kirkland (WA)

Hybrid
USD 90,000 - 150,000
Stock options
Gym membership
Health/dental/vision insurance
+4
On-Device Speech ML Researcher
On-Device Speech ML Researcher

Apple Inc. • Cupertino (CA), Northern (KY)

Hybrid
USD 150,000 - 225,000
Medical and dental coverage
Retirement benefits
Apple stock programs
+2
Speech & Audio ML Research Engineer
Speech & Audio ML Research Engineer

Hume AI • New York (NY)

On-site
USD 120,000 - 180,000
Human Factors Researcher: Trust & Usability in Social Computing
Human Factors Researcher: Trust & Usability in Social Computing

HushOne, Inc. • Kirkland (WA)

Hybrid
USD 100,000 - 170,000
Stock options
401(k) match
Health, dental, vision
+2
On-Device Multimodal Reasoning Architect
On-Device Multimodal Reasoning Architect

Apple Inc. • Sunnyvale (CA)

Hybrid
USD 150,000 - 278,000
Multimodal ML Researcher – On-Device Interactive AI
Multimodal ML Researcher – On-Device Interactive AI

Nuha • New York (NY)

On-site
USD 100,000 - 130,000
Doctoral Research Intern — Real-World, Collaborative Project
Doctoral Research Intern — Real-World, Collaborative Project

HushOne, Inc. • Kirkland (WA)

Hybrid
USD 90,000 - 110,000
Stock options
401(k) with company match
AI tokens
+1
Staff Multimodal Speech Engineer — Real-time AI
Staff Multimodal Speech Engineer — Real-time AI

Hark, Inc. • San Jose (CA)

On-site
USD 180,000 - 450,000
Staff Engineer, Speech & Audio for Multimodal AI
Staff Engineer, Speech & Audio for Multimodal AI

Hark • San Jose (CA)

On-site
USD 180,000 - 450,000
Lead ML Architect, Conversational Speech (On-Device)
Lead ML Architect, Conversational Speech (On-Device)

Apple Inc. • Cupertino (CA)

On-site
USD 263,000 - 394,000
Comprehensive medical and dental cover
Retirement benefits
Discounted products and free services
+2