Senior Machine Learning Engineer, Voice Agents

Jobgether

Netherlands

Hybrid

EUR 110,000 - 160,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Remote work options
Health, dental, and vision benefits
Parental leave and flexible PTO
Company equity
Conferences, training and education

Job summary

Jobgether is seeking a Senior Machine Learning Engineer for our partner company in the Netherlands to own substantial parts of an open voice-agent stack. You will drive architecture for a production-ready voice platform, integrating ASR, TTS, and end-to-end speech models while maintaining reliability and scalable APIs.

You will lead realtime GPU inference serving, streaming protocols, observability, and cross-team collaboration to deliver robust voice agents for developers and robotics

Qualifications

  • Senior-level engineering with ability to own a substantial architectural area.
  • Experience building developer-facing infrastructure in AI or ML environments.
  • Strong open-source contributions to a Python library and asynchronous Python proficiency.
  • Solid understanding of distributed systems and failure modes.
  • Proven shipping of real-time tech involving streaming, WebSockets/WebRTC, audio or live inference.
  • Hands-on production experience with LLMs or multimodal models.
  • Excellent written communication and collaborative skills, publicly.
  • Genuine interest in voice technology and conversational AI.

Responsibilities

  • Take architectural ownership of major parts of the open-source speech-to-speech library and pipelines.
  • Integrate new ASR, TTS, and end-to-end speech models while keeping clean abstractions.
  • Review contributions, triage issues, manage releases, grow the contributor community.
  • Design developer API and streaming protocol for the voice platform (WebSockets, session lifecycle, authentication).
  • Build and operate serving infra for realtime GPU inference with autoscaling and observability.
  • Collaborate with hub and inference teams to enable easy integration into products and demos.
  • Move platform from prototype to production with load testing and SLOs, degrade gracefully if needed.
  • Create docs, examples, and templates for quick onboarding to running voice agents.
  • Support deployments including robotics fleet and ongoing public tech content.
  • Contribute to the wider ML/AI community through talks and demos.

Skills

Open-source ownership
Asynchronous Python
Distributed systems
Realtime streaming
LLMs / multimodal models
Dev infra for AI
Public collaboration
Voice technology

Tools

WebSockets
WebRTC
GPU inference

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Machine Learning Engineer, Voice Agents based inNetherlands.

Own a major part of an open voice-agent stack at the intersection of machine learning, realtime systems, and developer infrastructure.
You will lead the architecture and evolution of an open-source speech-to-speech library while helping turn a new voice platform into a production-ready developer product.
The role combines deep backend and ML engineering with realtime audio, inference infrastructure, and developer experience.
You will integrate rapidly evolving ASR, TTS, and end-to-end speech models while maintaining clean abstractions and strong reliability.
You will have significant autonomy to shape APIs, streaming protocols, GPU serving, observability, and production architecture.
Your work will directly enable developers to build and deploy sophisticated voice agents and will support existing realtime deployments, including robotics applications.
This is an open, highly collaborative environment where you can contribute publicly through documentation, demos, talks, and open-source development.

Accountabilities
  • Take architectural ownership of significant parts of the open-source speech-to-speech library, including pipeline design, latency budgets, and realtime loop reliability.
  • Integrate new ASR, TTS, and end-to-end speech models as they become available while maintaining clean, extensible abstractions.
  • Review community contributions, triage issues, manage releases, and help grow the contributor community around the project.
  • Design the developer API and streaming protocol for the voice platform, including session lifecycle, WebSockets/WebRTC transport, authentication, error semantics, and versioning.
  • Build and operate the serving infrastructure for realtime GPU inference, including concurrency, autoscaling, observability, and cost-per-session optimization.
  • Collaborate with Hub and inference teams to make voice agents easy to integrate into products, applications, and demonstrations.
  • Take the platform from prototype to production through load testing, SLO definition, reliability improvements, and graceful degradation when models or network paths fail.
  • Create documentation, examples, and templates that enable developers to move from initial setup to a running voice agent quickly.
  • Support deployments already relying on the technology, including the existing robotics fleet.
  • Contribute to the wider technical community through blog posts, demonstrations, conference talks, or other public technical content when desired.
Requirements
  • Senior-level engineering experience with the ability to independently own a substantial part of an architecture and drive it forward.
  • Experience building developer-facing infrastructure in AI, machine learning, developer tools, or a comparable technical environment, such as inference APIs or agent infrastructure.
  • Significant open-source contributions to a Python library and strong proficiency with asynchronous Python.
  • Solid understanding of distributed systems and their failure modes.
  • Proven experience shipping realtime technology involving streaming, WebSockets, WebRTC, audio or video pipelines, or live inference.
  • Practical production experience with LLMs or multimodal models.
  • Strong written communication skills and a demonstrated ability to collaborate asynchronously and in public.
  • Genuine interest in voice technology and conversational AI.
  • Contributions to voice-agent frameworks such as speech-to-speech, Pipecat, LiveKit Agents, Vocode, or TEN are a plus.
  • Experience contributing to llama.cpp or another low-level inference runtime is advantageous.
  • Hands-on experience with ASR, TTS, or end-to-end speech models, including evaluating latency and quality trade-offs, is a plus.
  • GPU serving, quantization, or on-device inference experience is advantageous.
  • Knowledge of audio pipelines, including VAD, echo cancellation, jitter buffers, barge-in, and turn detection, is a plus.
  • Experience deploying technology to embedded or robotics environments is beneficial.
  • A public technical track record through talks, blog posts, demos, or similar contributions is valued.
  • Candidates are encouraged to apply even if they do not meet every listed requirement, particularly where their experience could bring complementary strengths to the team.
Benefits
  • Flexible working hours and remote work options.
  • Health, dental, and vision benefits for employees and their dependents.
  • Parental leave and flexible paid time off.
  • Company equity as part of the compensation package for all employees.
  • Reimbursement for relevant conferences, training, and education to support continuous professional development.
  • Access to a distributed, international work environment with opportunities to collaborate with experienced professionals across the AI and machine learning community.
  • Opportunities for remote employees to visit company offices in New York City and Paris.
  • Workstation equipment and setup support when needed to help employees work effectively.
  • A strong commitment to diversity, equity, and inclusion, with a workplace designed to ensure employees feel respected and supported.
  • Opportunities to contribute to and connect with the broader ML/AI community.
  • A culture focused on impact, continuous learning, collaboration, and professional growth.

We appreciate your interest and wish you the best!

Data Privacy Notice: By submitting your application, you acknowledge that Jobgether will process your personal data to evaluate your candidacy and share relevant information with the hiring employer. This processing is based on legitimate interest and pre-contractual measures under applicable data protection laws (including GDPR). You may exercise your rights (access, rectification, erasure, objection) at any time.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior ML Engineer, Voice Agents - Remote, Equity
Senior ML Engineer, Voice Agents - Remote, Equity

Jobgether • Netherlands

Hybrid
EUR 110,000 - 160,000
Remote work options
Health, dental, and vision benefits
Parental leave and flexible PTO
+2
Software Engineer, Data Infrastructure & Acquisition - Utrecht, Netherlands
Software Engineer, Data Infrastructure & Acquisition - Utrecht, Netherlands

Speechify • Utrecht

On-site
EUR 26,000 - 103,000
Fast-growing environment
Entrepreneurial-minded team
Hands-off management
+3
Senior Software Engineer, Core Experiences - Amsterdam, Netherlands
Senior Software Engineer, Core Experiences - Amsterdam, Netherlands

Clutch Canada • Amsterdam

On-site
EUR 70,000 - 90,000
Competitive salaries
Friendly and laid-back atmosphere
Opportunity for significant impact
+2
Senior Software Engineer, Core Experiences - Eindhoven, Netherlands
Senior Software Engineer, Core Experiences - Eindhoven, Netherlands

Clutch Canada • Eindhoven

On-site
EUR 55,000 - 75,000
Competitive salaries
Friendly and laid-back atmosphere
Hands-off management approach
Enterprise Solutions Engineer - Netherlands
Enterprise Solutions Engineer - Netherlands

ElevenLabs • Netherlands

On-site
EUR 60,000 - 80,000
Innovative culture
Growth paths
Learning & development stipend
+3
Senior Machine Learning Engineer – LLMs
Senior Machine Learning Engineer – LLMs

Prosus • Amsterdam

Hybrid
EUR 120,000 - 190,000
High-impact AI projects
Expert colleagues
Competitive compensation
+1
Senior Software Engineer, Core Experiences - Utrecht, Netherlands
Senior Software Engineer, Core Experiences - Utrecht, Netherlands

Speechify • Utrecht

On-site
EUR 26,000 - 104,000
Competitive salary
Stock options
Remote-friendly culture
+2
Senior Software Engineer, Core Experiences - Rotterdam, Netherlands
Senior Software Engineer, Core Experiences - Rotterdam, Netherlands

Speechify • Rotterdam

On-site
EUR 55,000 - 85,000
Competitive salaries
Friendly and laid-back atmosphere
Opportunity for impact in a growing industry
Senior Machine Learning Engineer (AI)
Senior Machine Learning Engineer (AI)

Bond Personnel Group • Rotterdam, Amsterdam

On-site
EUR 65,000 - 85,000
Flexible working hours
Dedicated time for passion projects
Supportive, learning-focused startup environment
Software Engineer, Platform - The Hague, Netherlands
Software Engineer, Platform - The Hague, Netherlands

Speechify • Den Haag

On-site
EUR 50,000 - 80,000
Dynamic work environment
Competitive compensation
Autonomy in work