Real-Time Speech ML Engineer — Sub-Second Latency

Sellsig

Minnesota

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Sellsig is seeking an experienced ML Engineer to build real-time speech and language processing pipelines. The role focuses on developing streaming ASR and objection detection systems.

With at least 4 years of experience in applied ML, candidates should excel in Python and have practical exposure to low-latency implementations. The position emphasizes driving end-to-end latency and integrating modern LLM tools.

Qualifications

  • 4+ years in ML/applied ML, with production work in speech, NLP, or real-time inference.
  • Experience deploying models in low-latency, streaming settings.
  • Comfortable owning ambiguous, research-flavored problems end to end.

Responsibilities

  • Build and tune streaming ASR and intent/objection-detection pipelines.
  • Drive down end-to-end latency while holding accuracy.
  • Evaluate and integrate speech and LLM providers.

Skills

Machine Learning
Python
ASR/diarization
Low-latency streaming
Evaluation metrics
WebRTC

Tools

Deepgram
AssemblyAI

Job description

Sellsig is seeking an experienced ML Engineer to build real-time speech and language processing pipelines. The role focuses on developing streaming ASR and objection detection systems.

With at least 4 years of experience in applied ML, candidates should excel in Python and have practical exposure to low-latency implementations. The position emphasizes driving end-to-end latency and integrating modern LLM tools.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Engineer — Real-time Speech
ML Engineer — Real-time Speech

Sellsig • Minnesota

On-site
USD 100,000 - 130,000
Edge ML Engineer: Real-Time Speech & Language for Autonomy
Edge ML Engineer: Real-Time Speech & Language for Autonomy

Instant Teams • New Haven (CT)

On-site
USD 150,000 - 215,000
Equity stake
100% Remote (US)
Comprehensive Health Coverage
+3
Senior Speech ML Engineer — Scalable ASR & Real‑Time Insights
Senior Speech ML Engineer — Scalable ASR & Real‑Time Insights

Level AI • Mountain View (CA)

On-site
USD 120,000 - 170,000
ML Engineer: Real-Time Audio & Production ML
ML Engineer: Real-Time Audio & Production ML

Catalyst Labs • Santa Clara (CA)

On-site
USD 80,000 - 120,000
Founding ML Engineer: Generalizable Subvocal Speech
Founding ML Engineer: Generalizable Subvocal Speech

Socket.dev • San Francisco (CA)

On-site
USD 180,000 - 250,000
ML Engineer – Real-Time Audio AI & Production Systems
ML Engineer – Real-Time Audio AI & Production Systems

Catalyst Labs • Austin (TX)

On-site
USD 90,000 - 130,000
ML Engineer: Real-Time Audio AI for Field Sales
ML Engineer: Real-Time Audio AI for Field Sales

Catalyst Labs • Houston (TX)

On-site
USD 120,000 - 180,000
ML Engineer – Real-Time Audio AI & Production Systems
ML Engineer – Real-Time Audio AI & Production Systems

Catalyst Labs • San Jose (CA)

On-site
USD 100,000 - 140,000
Real-Time Voice AI Engineer (ML/LLM) — Equity
Real-Time Voice AI Engineer (ML/LLM) — Equity

Mia • Austin (TX)

Hybrid
USD 110,000 - 170,000
Equity (stock options)
Health, vision & dental insurance
Flexible PTO
+3
ML Engineer
ML Engineer

Catalyst Labs • San Jose (CA)

On-site
USD 100,000 - 140,000