AI/LLM Speech & Audio AI Evaluation Specialist (Int. Voice)

Mindtel

Dadri

On-site

INR 900,000 - 1,300,000

Full time

4 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

On-site innovation lab in India
Flexible hours

Job summary

Mindtel is seeking an interpreter for an on-site role in India to evaluate and benchmark LLM-driven speech outputs across fluency, naturalness, and contextual accuracy. The candidate will design A/B tests, annotate datasets, and collaborate with ML engineers to inform model fine-tuning, with a focus on multilingual voice applications for fintech, healthcare, and telco sectors.

The role requires strong Python skills, experience with speech-to-text and text-to-speech systems, and familiarity with

Qualifications

  • Proficiency in Python and ASR/TTS pipelines.
  • Experience evaluating LLM-driven speech outputs using MOS/WER/CER.
  • Experience annotating and labeling speech datasets for model training.
  • Familiarity with multilingual speech and voice UX.

Responsibilities

  • Evaluate and benchmark LLM-driven speech outputs across fluency, naturalness, emotional tone, and contextual accuracy using standardized metrics and custom rubrics.
  • Design and run A/B tests for voice models across accents, registers, and cultural contexts to improve perceived quality and user trust.
  • Annotate, transcribe, and label voice datasets for training and evaluation - focusing on turn-taking, intent alignment, and paralinguistic features.
  • Collaborate with ML engineers to translate evaluation insights into fine-tuning signals and data augmentation strategies.
  • Document voice quality trends, failure modes, and improvement pathways in structured reports for cross-functional review.
  • Stay current with research in speech AI, prosody, and LLM evaluation - applying new methods like pairwise preference ranking or automated metrics (e.g., MOS, WER, CER).

Skills

Python
Speech-to-Text
Text-to-Speech
Speech Evaluation Metrics
LLM Evaluation
Audio Annotation Tools
Conversational AI
Multilingual Speech
Prosody Analysis
Paralinguistic Features
Voice UX Design

Job description

About the Opportunity

A fast-growing AI innovation lab focused on next-gen voice and speech intelligence, we build and evaluate cutting-edge Large Language Models (LLMs) for real-world conversational AI applications. Our work spans multilingual voice agents, prosody modeling, and context-aware synthesis - serving global clients across fintech, healthcare, and telco sectors. We move fast, validate rigorously, and ship voice experiences that sound human, behave intelligently, and scale reliably.

Role & Responsibilities
  • Evaluate and benchmark LLM-driven speech outputs across fluency, naturalness, emotional tone, and contextual accuracy using standardized metrics and custom rubrics.
  • Design and run A/B tests for voice models across accents, registers, and cultural contexts to improve perceived quality and user trust.
  • Annotate, transcribe, and label voice datasets for training and evaluation - focusing on turn-taking, intent alignment, and paralinguistic features.
  • Collaborate with ML engineers to translate evaluation insights into fine-tuning signals and data augmentation strategies.
  • Document voice quality trends, failure modes, and improvement pathways in structured reports for cross-functional review.
  • Stay current with research in speech AI, prosody, and LLM evaluation - applying new methods like pairwise preference ranking or automated metrics (e.g., MOS, WER, CER).
Skills & Qualifications
  • Must-Have
  • Python
  • Speech-to-Text
  • Text-to-Speech
  • Speech Evaluation Metrics
  • LLM Evaluation
  • Audio Annotation Tools
  • Conversational AI
  • Multilingual Speech
  • Preferred
  • Prosody Analysis
  • Paralinguistic Features
  • Voice UX Design
Benefits & Culture Highlights
  • Work on bleeding-edge voice AI projects with global impact.
  • Collaborative, research-driven culture with access to top-tier LLM tooling.
  • On-site innovation lab environment in India with flexible hours and peer learning.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Speech & Audio AI Evaluation Specialist (International Voice)
Speech & Audio AI Evaluation Specialist (International Voice)

Innodata India • Dadri

On-site
INR 900,000 - 1,500,000
Applied AI Engineer
Applied AI Engineer

Synth (YC S21) • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Analyst (Speech and Audio AI evaluation)
Analyst (Speech and Audio AI evaluation)

Innodata Inc. • Dadri

On-site
INR 900,000 - 1,350,000
Lead AI Engineer – Agentic Systems & Voice AI
Lead AI Engineer – Agentic Systems & Voice AI

YAL • Hyderabad

On-site
INR 4,500,000 - 9,000,000
ML Researcher — Speech
ML Researcher — Speech

RippleWorks Inc. • India

On-site
INR 1,200,000 - 1,800,000
AI/ML Technical Lead – Conversational Intelligence 5-12 Years Exp – Kolkata
AI/ML Technical Lead – Conversational Intelligence 5-12 Years Exp – Kolkata

HypTechie • Kolkata District

On-site
INR 4,000,000 - 6,000,000
Speech & Audio AI Specialist (from int. bpo only)
Speech & Audio AI Specialist (from int. bpo only)

Beyond Border Consultants • Dadri

On-site
INR 1,400,000 - 2,200,000
On-site work
GPU clusters access
Global team exposure
ML Researcher — Speech
ML Researcher — Speech

Rippleworks • India

On-site
INR 1,200,000 - 2,800,000
AI Trainer
AI Trainer

Innodata India • India

On-site
INR 1,004,000 - 1,339,000
Technical Project Manager
Technical Project Manager

Bigship • Dehradun

On-site
INR 1,800,000 - 3,000,000
GPU compute resources
AI research budget
Flexible working arrangements