Audio Inference Engineer — Model Efficiency (Remote-friendly)

Cohere

New York (NY)

Hybrid

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
Parental Leave top-up for 6 months
Personal enrichment benefits
6 weeks of vacation

Job summary

Cohere is seeking an engineering professional in New York to develop and optimize audio machine learning systems. You will work with cross-functional teams to improve audio model metrics, addressing latency and throughput while ensuring real-time audio inference integration. The ideal candidate will possess expertise in programming languages like C++ and Python, and experience with audio deep learning models. Cohere offers a vibrant workplace with a focus on AI research, full health benefits, six weeks of vacation, and an inclusive culture.

Qualifications

  • Significant experience developing high-performance audio or machine learning inference systems.
  • Proficiency with C++ and Python.
  • Hands-on experience with deep learning models for audio, speech, or language applications.

Responsibilities

  • Work on advancing core audio model serving metrics, including latency, throughput, and quality.
  • Identify bottlenecks and deliver creative solutions for audio processing and streaming workloads.
  • Collaborate with training and serving infrastructure teams for seamless integration.

Skills

High-performance audio or machine learning inference systems
C++
Python
Deep learning models for audio
Results-oriented mindset

Tools

PyTorch
TensorFlow
vLLM
SGLang
Tensort-LLM

Job description

Cohere is seeking an engineering professional in New York to develop and optimize audio machine learning systems. You will work with cross-functional teams to improve audio model metrics, addressing latency and throughput while ensuring real-time audio inference integration. The ideal candidate will possess expertise in programming languages like C++ and Python, and experience with audio deep learning models. Cohere offers a vibrant workplace with a focus on AI research, full health benefits, six weeks of vacation, and an inclusive culture.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Audio Inference Engineer, Model Efficiency
Remote Audio Inference Engineer, Model Efficiency

Jaide Health • San Francisco (CA)

On-site
USD 100,000 - 140,000
Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
+2
Audio Inference Engineer, Model Efficiency
Audio Inference Engineer, Model Efficiency

Visa Hunt • New York (NY)

Hybrid
USD 140,000 - 190,000
Weekly lunch stipend
Health and dental benefits
RRSP matching / 401K / Pension
+6
Remote Audio Inference Engineer — Fast ML Serving
Remote Audio Inference Engineer — Fast ML Serving

Visa Hunt • New York (NY)

Hybrid
USD 140,000 - 190,000
Weekly lunch stipend
Health and dental benefits
RRSP matching / 401K / Pension
+6
Remote Audio Inference Engineer — High-Performance Serving
Remote Audio Inference Engineer — High-Performance Serving

Cohere • United States

Remote
USD 110,000 - 180,000
A weekly lunch stipend of $75/£75 or"
Audio Inference Engineer, Model Efficiency
Audio Inference Engineer, Model Efficiency

Cohere • New York (NY)

Hybrid
USD 120,000 - 150,000
Open and inclusive culture
Weekly lunch stipend
Full health and dental benefits
+3
Audio ML Engineer - Real-World Voice & On-Device
Audio ML Engineer - Real-World Voice & On-Device

Sandbar • New York (NY)

On-site
USD 170,000 - 240,000
Remote Research Engineer — AI Audio & Models
Remote Research Engineer — AI Audio & Models

ElevenLabs • Town of Poland (NY)

On-site
USD 140,000 - 190,000
Staff Engineer, Model Efficiency & LLM Inference
Staff Engineer, Model Efficiency & LLM Inference

Visa Hunt • New York (NY)

Hybrid
USD 150,000 - 210,000
Lunch stipend
Health and dental benefits
RRSP matching
+4
ML Engineer – Real-Time Audio AI & Production Systems
ML Engineer – Real-Time Audio AI & Production Systems

Catalyst Labs • New Jersey

On-site
USD 100,000 - 150,000
Audio AI Researcher: Multimodal, Low-Latency Modeling
Audio AI Researcher: Multimodal, Low-Latency Modeling

Mosaic.tech • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health benefits
Dental benefits
Vision benefits
+3