Senior Machine Learning Engineer, Voice Agents

Jobgether

France

Sur place

EUR 110 000 - 150 000

Plein temps

Il y a 4 jours
Soyez parmi les premiers à postuler

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Flexible hours
Remote work options
Health benefits
Parental leave
Equity
Conferences & training
Office visits NYC/Paris
Equipment setup support
Diversity & inclusion
ML/AI community

Résumé du poste

Jobgether is seeking a Senior Machine Learning Engineer, Voice Agents, based in France. You will own substantial parts of an open voice-agent stack, shaping APIs, streaming protocols, and GPU inference infrastructure for real-time speech systems.

You will integrate ASR, TTS, and end-to-end speech models, contribute to documentation, and help productionize the platform for robotics deployments and diverse demos.

Qualifications

  • Senior-level engineering experience with owning a substantial architecture.
  • Experience building developer-facing infrastructure in AI or ML environments.
  • Significant open-source contributions to a Python library and asynchronous Python.
  • Strong understanding of distributed systems and failure modes.
  • Shipping realtime tech involving streaming, WebSockets, WebRTC, audio/video pipelines.
  • Practical production experience with LLMs or multimodal models.
  • Strong written communication and async collaboration in public.
  • Genuine interest in voice technology and conversational AI.
  • Experience with voice-agent frameworks or low-level inference runtimes is a plus.

Responsabilités

  • Take architectural ownership of parts of the open-source speech-to-speech library.
  • Integrate new ASR, TTS, and end-to-end speech models with clean abstractions.
  • Review community contributions and manage releases to grow contributors.
  • Design developer API and streaming protocol for the voice platform.
  • Build and operate serving infra for realtime GPU inference with observability.
  • Collaborate with Hub and inference teams for easy integration into products.
  • Move platform from prototype to production with load testing and SLOs.
  • Create docs, examples, and templates for developers to run voice agents.
  • Support deployments including robotics fleet and public demonstrations.
  • Contribute to the broader ML/AI community via talks or content.

Connaissances

Python (async)
Distributed systems
Realtime streaming
Open-source contributions
Public collaboration
Voice tech interest

Outils

WebSockets
WebRTC
GPU serving
LLMs / AI inference
llama.cpp
Audio pipelines

Description du poste

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Machine Learning Engineer, Voice Agents based in France.

Own a major part of an open voice-agent stack at the intersection of machine learning, realtime systems, and developer infrastructure.
You will lead the architecture and evolution of an open-source speech-to-speech library while helping turn a new voice platform into a production-ready developer product.
The role combines deep backend and ML engineering with realtime audio, inference infrastructure, and developer experience.
You will integrate rapidly evolving ASR, TTS, and end-to-end speech models while maintaining clean abstractions and strong reliability.
You will have significant autonomy to shape APIs, streaming protocols, GPU serving, observability, and production architecture.
Your work will directly enable developers to build and deploy sophisticated voice agents and will support existing realtime deployments, including robotics applications.
This is an open, highly collaborative environment where you can contribute publicly through documentation, demos, talks, and open-source development.

Accountabilities
  • Take architectural ownership of significant parts of the open-source speech-to-speech library, including pipeline design, latency budgets, and realtime loop reliability.
  • Integrate new ASR, TTS, and end-to-end speech models as they become available while maintaining clean, extensible abstractions.
  • Review community contributions, triage issues, manage releases, and help grow the contributor community around the project.
  • Design the developer API and streaming protocol for the voice platform, including session lifecycle, WebSockets/WebRTC transport, authentication, error semantics, and versioning.
  • Build and operate the serving infrastructure for realtime GPU inference, including concurrency, autoscaling, observability, and cost-per-session optimization.
  • Collaborate with Hub and inference teams to make voice agents easy to integrate into products, applications, and demonstrations.
  • Take the platform from prototype to production through load testing, SLO definition, reliability improvements, and graceful degradation when models or network paths fail.
  • Create documentation, examples, and templates that enable developers to move from initial setup to a running voice agent quickly.
  • Support deployments already relying on the technology, including the existing robotics fleet.
  • Contribute to the wider technical community through blog posts, demonstrations, conference talks, or other public technical content when desired.
Requirements
  • Senior-level engineering experience with the ability to independently own a substantial part of an architecture and drive it forward.
  • Experience building developer-facing infrastructure in AI, machine learning, developer tools, or a comparable technical environment, such as inference APIs or agent infrastructure.
  • Significant open-source contributions to a Python library and strong proficiency with asynchronous Python.
  • Solid understanding of distributed systems and their failure modes.
  • Proven experience shipping realtime technology involving streaming, WebSockets, WebRTC, audio or video pipelines, or live inference.
  • Practical production experience with LLMs or multimodal models.
  • Strong written communication skills and a demonstrated ability to collaborate asynchronously and in public.
  • Genuine interest in voice technology and conversational AI.
  • Contributions to voice-agent frameworks such as speech-to-speech, Pipecat, LiveKit Agents, Vocode, or TEN are a plus.
  • Experience contributing to llama.cpp or another low-level inference runtime is advantageous.
  • Hands-on experience with ASR, TTS, or end-to-end speech models, including evaluating latency and quality trade-offs, is a plus.
  • GPU serving, quantization, or on-device inference experience is advantageous.
  • Knowledge of audio pipelines, including VAD, echo cancellation, jitter buffers, barge-in, and turn detection, is a plus.
  • Experience deploying technology to embedded or robotics environments is beneficial.
  • A public technical track record through talks, blog posts, demos, or similar contributions is valued.
  • Candidates are encouraged to apply even if they do not meet every listed requirement, particularly where their experience could bring complementary strengths to the team.
Benefits
  • Flexible working hours and remote work options.
  • Health, dental, and vision benefits for employees and their dependents.
  • Parental leave and flexible paid time off.
  • Company equity as part of the compensation package for all employees.
  • Reimbursement for relevant conferences, training, and education to support continuous professional development.
  • Access to a distributed, international work environment with opportunities to collaborate with experienced professionals across the AI and machine learning community.
  • Opportunities for remote employees to visit company offices in New York City and Paris.
  • Workstation equipment and setup support when needed to help employees work effectively.
  • A strong commitment to diversity, equity, and inclusion, with a workplace designed to ensure employees feel respected and supported.
  • Opportunities to contribute to and connect with the broader ML/AI community.
  • A culture focused on impact, continuous learning, collaboration, and professional growth.

We appreciate your interest and wish you the best!

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Senior ML Engineer, Voice Agents — Remote & Production-Ready
Senior ML Engineer, Voice Agents — Remote & Production-Ready

Jobgether • France

Sur place
EUR 110 000 - 150 000
Flexible hours
Remote work options
Health benefits
+7
Senior AI Software Engineer
Senior AI Software Engineer

GetVocal AI Ltd. • Paris

Hybride
EUR 90 000 - 130 000
25 days holiday
Private healthcare
Diversified, international team
+1
Senior AI Engineer (m/f/d)
Senior AI Engineer (m/f/d)

Jobgether • France

Hybride
EUR 90 000 - 140 000
Home-office budget
Learning budget
Relocation support
+3
Senior AI Software Engineer
Senior AI Software Engineer

GetVocal AI • Paris

Sur place
EUR 110 000 - 160 000
High ownership and autonomy
Diverse, international team across*欧洲
Exposure to cutting-edge AI, voice, &
+4
Senior ML Manager - Voice Teams (x/f/m)
Senior ML Manager - Voice Teams (x/f/m)

Doctolib • Paris

Hybride
EUR 120 000 - 170 000
Full remote flexibility
Hybrid work model
Applied AI Engineer
Applied AI Engineer

Norbert Health • Paris

Sur place
EUR 70 000 - 90 000
Competitive salary and equity
High autonomy and technical ownership
Transparent, mission-driven culture
Enterprise Solutions Engineer - France
Enterprise Solutions Engineer - France

ElevenLabs • France

À distance
EUR 50 000 - 70 000
Professional development stipend
Annual company offsite
Co-working stipend
Software Engineer, Data Infrastructure & Acquisition - Lille, France
Software Engineer, Data Infrastructure & Acquisition - Lille, France

Clutch Canada • Lille

Sur place
EUR 68 000 - 103 000
Competitive salaries
Laid-back atmosphere
Opportunity to impact learning differences
Software Engineer, Data Infrastructure & Acquisition - Paris, France
Software Engineer, Data Infrastructure & Acquisition - Paris, France

Clutch Canada • Paris

Sur place
EUR 50 000 - 80 000
Competitive salaries
Flexible work environment
Friendly atmosphere
+1
Voice AI Engineer
Voice AI Engineer

exponentiel.ai • Paris

Sur place
EUR 60 000 - 95 000