Audio Modeling Engineer — Real-Time Production Inference

attention labs

San Francisco, Northern (CA, KY)

Hybrid

USD 150,000 - 200,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

attention labs is building the audio modeling stack in a fast-moving startup setting in San Francisco. You will own end-to-end lifecycle from data to production inference, shipping models that recognize who is being spoken to in real time.

You will turn research into dependable inference under real-world noise and latency constraints, design evaluation harnesses, and collaborate with founders to shape next steps for the product's audio intelligence.

Qualifications

  • Strong applied ML experience with audio, speech, or signal-heavy models.
  • Ability to own a model from data collection to production inference, not just notebooks.
  • Fluency with modern deep learning tooling such as PyTorch and rigorous measurement.

Responsibilities

  • Own the audio model lifecycle end to end: data, features, architecture, training, evaluation, and production inference.
  • Build and improve models behind addressee detection and real-time decisions on who is being spoken to.
  • Turn research prototypes into fast, dependable inference under real-world noise and accents.
  • Design evaluation harnesses and datasets to honestly measure impact of changes.
  • Work directly with founders and early customers to shape next model objectives.

Skills

Audio/speech ML
Production model lifecycle
PyTorch
Latency awareness

Tools

PyTorch

Job description

attention labs is building the audio modeling stack in a fast-moving startup setting in San Francisco. You will own end-to-end lifecycle from data to production inference, shipping models that recognize who is being spoken to in real time.

You will turn research into dependable inference under real-world noise and latency constraints, design evaluation harnesses, and collaborate with founders to shape next steps for the product's audio intelligence.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Engineer, Speech/Audio
Machine Learning Engineer, Speech/Audio

attention labs • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 200,000
Embedded Systems Engineer, On-Device Inference
Embedded Systems Engineer, On-Device Inference

attention labs • Memphis (TN), Northern (KY)

Hybrid
USD 110,000 - 150,000
Audio AI Researcher: Multimodal, Low-Latency Modeling
Audio AI Researcher: Multimodal, Low-Latency Modeling

Mosaic.tech • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3
ML Engineer: Real-Time Audio AI for Field Sales
ML Engineer: Real-Time Audio AI for Field Sales

Catalyst Labs • Houston (TX)

On-site
USD 120,000 - 180,000
Lead Real-Time Voice AI Architect
Lead Real-Time Voice AI Architect

Inflection AI, Inc. • Palo Alto (CA), Northern (KY)

Hybrid
USD 400,000 - 550,000
Robust medical, dental and vision with
401k matching
Flexible Time Off
+3
Remote Audio Inference Engineer — High-Performance Serving
Remote Audio Inference Engineer — High-Performance Serving

Cohere • United States

Remote
USD 110,000 - 180,000
A weekly lunch stipend of $75/£75 or"
Remote Audio Inference Engineer, Model Efficiency
Remote Audio Inference Engineer, Model Efficiency

Jaide Health • San Francisco (CA)

On-site
USD 100,000 - 140,000
Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
+2
Audio Inference Engineer — Model Efficiency (Remote-friendly)
Audio Inference Engineer — Model Efficiency (Remote-friendly)

Cohere • New York (NY)

Hybrid
USD 120,000 - 150,000
Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
+3
Senior Inference Engineer, AGI
Senior Inference Engineer, AGI

Amazon • Sunnyvale (CA)

On-site
USD 192,000 - 260,000
Health insurance
401(k) matching
RSU/Stock options
Research, Audio Expertise
Research, Audio Expertise

Mosaic.tech • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3