Senior ML Engineer - Text-to-Speech & Scalable Inference

ConnexAI

Manchester

On-site

GBP 90,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ConnexAI in Manchester, United Kingdom, is seeking a Senior Machine Learning Engineer to deploy state‑of‑the‑art Text‑to‑Speech models and scale TTS systems for production. You will optimise pipelines for GPU performance and ensure reliable inference at scale.

You will work with PyTorch and Hugging Face transformers, integrate LLM inference servers and complex production pipelines, and push the boundaries of neural audio codecs like Encodec.

Qualifications

  • MSc or PhD in Computer Science or a related field.
  • 3–5 years of hands‑on experience deploying and scaling ML solutions in production.
  • Proven experience deploying and optimising LLMs/transformers in production.
  • Knowledge of LLM inference servers (e.g., Triton, TensorRT, TorchServe).
  • Experience with GPU scaling for large‑scale ML models.
  • Expertise in deploying complex ML pipelines in production environments.
  • Proficiency with PyTorch and Hugging Face transformers.
  • Experience with neural audio codecs (e.g., Encodec).
  • Background in Text‑to‑Speech (TTS) development.
  • Experience with RVQ, GANs, and diffusion models.

Responsibilities

  • Collaborate closely with the TTS team to deploy and scale advanced models in production environments.
  • Lead efforts in optimising TTS pipelines for performance and scalability, particularly focusing on GPU utilisation.
  • Implement and maintain LLM and transformer models, ensuring efficient inference at scale.
  • Integrate and manage LLM‑based inference servers like Triton, TensorRT, or TorchServe to streamline deployment.
  • Work on deploying complex pipelines in production, ensuring seamless integration with existing systems.

Skills

LLM deployment
PyTorch
Hugging Face
GPU scaling
Transformer models
TTS development
Inference optimization
TorchServe/Triton
VPUs/ GPUs optimization

Education

MSc or PhD in Computer Science or related field

Tools

Triton
TensorRT
TorchServe

Job description

ConnexAI in Manchester, United Kingdom, is seeking a Senior Machine Learning Engineer to deploy state‑of‑the‑art Text‑to‑Speech models and scale TTS systems for production. You will optimise pipelines for GPU performance and ensure reliable inference at scale.

You will work with PyTorch and Hugging Face transformers, integrate LLM inference servers and complex production pipelines, and push the boundaries of neural audio codecs like Encodec.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Engineer
Machine Learning Engineer

ConnexAI • Manchester

On-site
GBP 90,000 - 140,000
Senior ML Engineer: Conversational AI, Edge & Multi-Cloud
Senior ML Engineer: Conversational AI, Edge & Multi-Cloud

Connect • City Of London

On-site
GBP 90,000 - 130,000
ML Engineer, TTS & Voice AI
ML Engineer, TTS & Voice AI

Cantina Labs • Greater London

On-site
GBP 148,000 - 164,000
Equity
Health insurance
PTO 42 days
+3
Platform Backend Engineer — Scale TTS APIs & AI Audio
Platform Backend Engineer — Scale TTS APIs & AI Audio

Speechify • Reading

On-site
GBP 90,000 - 120,000
Remote work
Competitive compensation
Senior Software Engineer: Real-Time AI SaaS Leader
Senior Software Engineer: Real-Time AI SaaS Leader

ConnexAI • Manchester

Hybrid
GBP 70,000 - 110,000
ASR Research Scientist: Multilingual Real-Time Speech
ASR Research Scientist: Multilingual Real-Time Speech

ConnexAI • Manchester

On-site
GBP 70,000 - 120,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Connect • City Of London

On-site
GBP 90,000 - 130,000
Senior Audio AI Engineer: Remote TTS & Speech Synthesis
Senior Audio AI Engineer: Remote TTS & Speech Synthesis

Awarri • United Kingdom

Hybrid
GBP 60,000 - 80,000
Be part of a pioneering initiative
Work on impactful projects
Collaborate with a passionate team
Senior ML/AI Engineer - Hybrid, Production-Scale NLP & LLM
Senior ML/AI Engineer - Hybrid, Production-Scale NLP & LLM

Ocho People • Belfast City District

Hybrid
GBP 65,000 - 75,000
Hybrid work
Global exposure
High-impact ML work
Senior AI Engineer - Production-Ready Speech & NLP (Remote)
Senior AI Engineer - Production-Ready Speech & NLP (Remote)

Advanceai • United Kingdom

On-site
GBP 70,000 - 90,000
Competitive salary
Discretionary annual bonus
Full benefits package