AI Engineer, Agentic Voice (TTS)

United States Digital Space LLC

Kitchener

Hybrid

CAD 90,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

As an AI Engineer on our Speech Team, you’ll own the back-end implementation and linguistic optimization of the voice (TTS) layer for our next-generation AI agents.

You’ll work squarely within our Speech Team, a high-impact R&D and engineering group focused on speech recognition, enhancement, and synthesis; bridging core speech science and product engineering so our agents sound human, context-aware, and trustworthy.

Qualifications

  • 3+ years in Speech Synthesis (TTS) or applied speech ML.
  • Strong Python and PyTorch experience.

Responsibilities

  • TTS backend implementation: Own the integration and optimization of multiple TTS vendor APIs behind a unified interface with failover, and lead research and prototyping of open-source and in-house TTS architectures.
  • Latency & pipeline engineering: Minimize time-to-first-audio and end-to-end latency across the ASR → LLM → TTS pipeline while maintaining or optimizing voice quality, in partnership with ASR and Audio AI engineers.
  • Linguistic optimization: Apply your knowledge of phonetics and sociolinguistics to format TTS input for maximum naturalness — SSML tags, punctuation-driven prosody, and text normalization for names, numbers, dates, and currency.
  • Persona system & parameter exposure: Build the persona parameterization system and architect the logic that exposes voice attributes to the product UI, implementing the house standards defined by Agent Experience Design.
  • Prompt engineering as code: Manage structured LLM and TTS prompt templates with versioning and a rigorous evaluation harness.
  • Conversational turn design: Engineer context- and state-aware "thinking" utterances that maintain caller trust while tool calls and model steps run under the hood.

Skills

Python
PyTorch
Speech Synthesis

Job description

About the company

the company is the AI platform for customer experience, built to resolve customer problems in real time across voice and digital. Our AI agents learn from your best human agents and improve with every interaction, helping organizations understand their customers, deliver better experiences, increase operational efficiencies, and build a lasting competitive advantage.

Unlike legacy systems built to route and answer, or standalone agentic bot vendors built to deflect, the company was built to resolve. Our AI agents and human agents operate on a single platform with shared context, allowing Agentic AI to resolve issues, advance deals, and eliminate busywork through automation while seamlessly handing conversations to humans when needed, with full context preserved.

Market-leading brands, including Randstad, Motorola Solutions, Netflix, the San Diego Padres, the Colorado Rockies Baseball Club, and Cal Athletics, trust the company. the company is backed by Andreessen Horowitz, GV, ICONIQ Capital, and T-Mobile.

Being a Dialer

At the company, AI isn’t just a feature; it’s how our teams do their best work every day. We put powerful AI tools in every employee’s hands so they can move faster, think bigger, and achieve more.

We believe every conversation matters. And we’ve built the platform that turns those conversations into insight and action, for our customers and ourselves.

We look for people who are intensely curious and hold themselves to a high bar. Our ambition is significant, and achieving it requires a team that operates at the highest level. We seek individuals who embody our core traits: Scrappy, Curious, Optimistic, Persistent, and Empathetic.

Your role

As an AI Engineer on our Speech Team, you'll own the back-end implementation and linguistic optimization of the voice (TTS) layer for our next-generation AI agents. You'll work squarely within our Speech Team, a high-impact R&D and engineering group focused on speech recognition, enhancement, and synthesis; bridging core speech science and product engineering so our agents sound human, context-aware, and trustworthy.

You’ll build the systems that render voice: integrating and optimizing TTS engines, engineering the persona and parameter machinery, and exposing voice attributes to our customer-facing UI. You'll partner closely with our Voice Experience Designer, who authors and owns the persona standards and quality bar that you implement — you own the platform that makes their designs real, fast, and consistent at scale.

This position reports to our Senior Manager, AI Speech, is based at our Kitchener hub, and operates on a hybrid schedule.

What you’ll do
  • TTS backend implementation: Own the integration and optimization of multiple TTS vendor APIs behind a unified interface with failover, and lead research and prototyping of open-source and in-house TTS architectures.
  • Latency & pipeline engineering: Minimize time-to-first-audio and end-to-end latency across the ASR → LLM → TTS pipeline while maintaining or optimizing voice quality, in partnership with ASR and Audio AI engineers.
  • Linguistic optimization: Apply your knowledge of phonetics and sociolinguistics to format TTS input for maximum naturalness — SSML tags, punctuation-driven prosody, and text normalization for names, numbers, dates, and currency.
  • Persona system & parameter exposure: Build the persona parameterization system and architect the logic that exposes voice attributes to the product UI, implementing the house standards defined by Agent Experience Design.
  • Prompt engineering as code: Manage structured LLM and TTS prompt templates with versioning and a rigorous evaluation harness.
  • Conversational turn design: Engineer context- and state-aware "thinking" utterances that maintain caller trust while tool calls and model steps run under the hood.
Skills you’ll bring
  • Technical foundation: Strong Python and hands-on experience with deep learning frameworks (e.g. PyTorch).
  • Speech expertise: 3+ years in Speech Synthesis (TTS) or applied speech ML, including <
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Engineer, Agentic Voice (TTS)
AI Engineer, Agentic Voice (TTS)

United States Digital Space LLC • Vancouver

Hybrid
CAD 110,000 - 150,000
AI Engineer, Agentic Voice (TTS)
AI Engineer, Agentic Voice (TTS)

Dialpad Japan • Kitchener

Hybrid
CAD 145,000 - 173,000
AI Engineer, Agentic Voice (TTS)
AI Engineer, Agentic Voice (TTS)

Dialpad • Kitchener

Hybrid
CAD 145,000 - 173,000
Competitive salary
Comprehensive benefits
Growth opportunities
AI Engineer, Agentic Voice (TTS)
AI Engineer, Agentic Voice (TTS)

Dialpad • Vancouver

Hybrid
CAD 161,000 - 192,000
AI Engineer
AI Engineer

Dialpad • Vancouver

Hybrid
CAD 161,000 - 192,000
Comprehensive benefits
Competitive salary
Opportunities for growth
AI Engineer
AI Engineer

Dialpad • Kitchener

Hybrid
CAD 145,000 - 173,000
Competitive salary
Comprehensive benefits
Training programs
+2
AI Engineer, Agentic Voice (TTS)
AI Engineer, Agentic Voice (TTS)

Dialpad Japan • Vancouver

Hybrid
CAD 161,000 - 192,000
Competitive salary
Comprehensive benefits
Career growth
Applied Scientist
Applied Scientist

United States Digital Space LLC • Kitchener

On-site
CAD 110,000 - 150,000
Sr. Technical Support Engineer
Sr. Technical Support Engineer

United States Digital Space LLC • Kitchener

On-site
CAD 70,000 - 100,000
Sr. Software Engineer (Connect - Core App & Admin)
Sr. Software Engineer (Connect - Core App & Admin)

United States Digital Space LLC • Vancouver

On-site
CAD 120,000 - 160,000