Machine Learning Engineer Agent Intelligence And Evaluations

ixigo

New Delhi

On-site

INR 2,000,000 - 5,000,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

ESOPs
Competitive salary

Job summary

ixigo is seeking a full-time research fellowship focused on self-healing voice agents for enterprise support. You will design audio-native evaluation metrics, generate robust datasets across accents, and define LLM-based rubrics for task completion and recovery from tool failures.

The role emphasizes Python-based experimentation, familiarity with Whisper/Conformer, and observability stacks. PhD students in ML/NLP/speech are strongly preferred, with ESOPs and a competitive salary offered.

Qualifications

  • PhD in ML/NLP/Speech or related field; strong research record preferred.
  • Experience with speech and dialogue evaluation research; publications beneficial.
  • Proficiency in Python and building evaluation pipelines for audio/voice systems.

Responsibilities

  • Design evaluation frameworks for voice agents, including metrics for accuracy, latency, and empathy.
  • Build end-to-end observability for audio interactions linking STT, LLM, and TTS traces.
  • Develop self-improvement loops: mine production traces, generate fine-tuning data, validate fixes.

Skills

Python
LLM tool-use
Observability stacks
Speech models

Education

PhD in ML/NLP/Speech

Tools

Whisper
Conformer
OpenTelemetry
Langfuse
Arize
Hamming

Job description

Job Description:

Company Description

We're building self-healing voice agents for enterprise customer support within ixigo. The system has to know when it's failing, why it's failing, and how to fix itself before a human notices. This fellowship sits at the intelligence layer behind that work.

Job Description

Voice agents fail in ways traditional software doesn't. An ASR confidence drop on a regional accent misfires a tool call, an LLM hallucinates a policy because upstream latency broke turn-taking, and support teams roll these agents back within a week without anyone able to explain what went wrong.

What you'll work on

Evaluation frameworks. Text-only evals miss most of what matters in voice: barge-in, prosody, latency-induced errors, cross-turn context loss. You'll design audio-native metrics, generate adversarial conversational datasets across accents and edge cases, and build LLM-as-judge rubrics for task completion, empathy, and recovery from tool failures.

End-to-end observability. Tracing a failed interaction means correlating audio packets, STT hypotheses, LLM reasoning traces, tool calls, and TTS output back to a single conversation ID. You'll help shape the schema and analysis layer that makes cascade failures visible across the stack.

Self-improvement systems. Once you can measure and trace, the interesting work is closing the loop: mining production traces for failure patterns, generating targeted fine-tuning data or prompt updates, and validating that fixes hold under adversarial replay.

Who we're looking for

Someone who cares about the research questions for their own sake, and equally cares whether the work ships. Papers at Interspeech, ACL, NeurIPS, or EMNLP on speech, dialogue systems, agent evaluation, or human-AI interaction are directly relevant.

Comfortable in Python, and familiar with at least one of: speech models (Whisper, Conformer variants), LLM tool-use and agent frameworks, or observability stacks (OpenTelemetry, Langfuse, Arize, Hamming).

Current PhD students in ML, NLP, or speech are the strong default; exceptional MS students or research engineers with a publication track record are welcome to apply.

Nice to have

Prior work on evaluation methodology, dataset synthesis, or interpretability. Experience with real-time systems, telephony, or streaming pipelines. A blog, repo, or workshop paper that shows how you think in public.

This is a full-time role with a competitive salary and ESOPs.

Additional Information

Our Culture:ixigo is proud to have built an entrepreneurial culture that has become a folk-lore in the startup ecosystem. One in every four ixigems has gone on to build successful startups and companies. Our cultural values of integrity, empathy, ingenuity, awesomeness, and resilience have stood the tests of time and we’ve built a fun, flexible and creative work environment that is driven by people with a high degree of ownership. You will get to work with some of the smartest folks in the Indian startup ecosystem, and solve some of the toughest problems for the next billion users by using bleeding-edge technologies. Oh, and we have an awesome "play" area, great chai/coffee, free lunches (yes, they exist!) and a workspace you will fall in love with.

Requirements:

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer — Agent Intelligence & Evaluation
Research Engineer — Agent Intelligence & Evaluation

ixigo • Gurugram District

On-site
INR 900,000 - 1,400,000
Equity
Autonomy over tooling
Governance-friendly environment
Enterprise Solutions Engineer - India
Enterprise Solutions Engineer - India

ElevenLabs • India

Remote
INR 800,000 - 1,500,000
Innovative culture
Growth paths
Learning & development stipend
+3
Applied AI Scientist
Applied AI Scientist

sirrus.ai • Mumbai

On-site
INR 1,200,000 - 2,200,000
Mumbai office
Birthday & anniversary leaves
AI/ML Technical Lead – Conversational Intelligence 5-12 Years Exp – Kolkata
AI/ML Technical Lead – Conversational Intelligence 5-12 Years Exp – Kolkata

HypTechie • Kolkata District

On-site
INR 4,000,000 - 6,000,000
Flexible, collaborative work environment
Access to cutting-edge AI tools
Competitive salary and performance-based rewards
Applied AI Engineer
Applied AI Engineer

Infer • Karnataka

On-site
INR 1,500,000 - 2,500,000
AI/ML Engineer
AI/ML Engineer

Cyfuture • Dadri

On-site
INR 900,000 - 1,200,000
Member of Technical Staff - Applied AI Research
Member of Technical Staff - Applied AI Research

Vaya • Delhi

On-site
INR 1,500,000 - 2,500,000
Machine Learning Engineer
Machine Learning Engineer

SourcingXPress • Gurugram District

On-site
INR 1,500,000 - 1,800,000
Member of Technical Staff - Applied AI Research
Member of Technical Staff - Applied AI Research

Anuvaya Labs • New Delhi

On-site
INR 1,200,000 - 1,800,000
Research Engineer Poland +4 more
Research Engineer Poland +4 more

ElevenLabs • Bengaluru

Remote
PLN 169,000 - 255,000
Innovative culture
Growth paths
Learning & development stipend
+2