Research Scientist - Speech/Audio Machine Learning

Institute of Foundation Models

Paris

Sur place

EUR 50 000 - 70 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

A leading AI research institution in France is seeking a Research Scientist specializing in speech/audio machine learning. You will design and train state-of-the-art neural speech models, focusing on innovative architectures for low-latency speech translation. A PhD or MSc in Computer Science, with publications in top venues, is essential. This role offers a unique opportunity to influence the development of transformative AI solutions that shape industries.

Qualifications

  • PhD or MSc focusing on Deep Learning, Signal Processing, or Computational Linguistics.
  • Published research in top-tier venues (NeurIPS, ICLR, ICASSP, Interspeech).

Responsabilités

  • Develop novel neural architectures for low-latency speech-to-speech translation.
  • Design custom objective functions to optimize prosody.
  • Conduct large-scale training runs and ablation studies.
  • Establish rigorous benchmarks using objective and subjective testing.

Connaissances

Deep Learning
Signal Processing
Computational Linguistics

Formation

PhD or MSc in Computer Science

Description du poste

About the Institute of Foundation Models

We are a dedicated research lab for building, understanding, using, and risk-managing foundation models. Our mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-driven economy.

As part of our team, you’ll have the opportunity to work on the core of cutting-edge foundation model training, alongside world-class researchers, data scientists, and engineers, tackling the most fundamental and impactful challenges in AI development. You will participate in the development of groundbreaking AI solutions that have the potential to reshape entire industries. Strategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers.

The Role

As a Research Scientist specializing in speech/audio machine learning, you will contribute tothe design and training ofSoTAend-to-end neural speech models. You will be responsible for developing the core intellectual property, moving beyond cascaded ASR → TTS systems toward native audio-to-audio multimodal architectures.

Key responsibilities
  • Architectural Design: Develop novel neural architectures for low-latency speech-to-speech translation and generation (e.g., Diffusion, Flow-matching, Transformer-based audio LLMs).
  • Loss Function Engineering: Design and implement customobjectivefunctions to optimize prosody(emotions, intelligibility, naturalness).
  • Experimental Iteration: Conduct large-scale training runs, performing ablation studies onmodel architectureand tokenization strategies.
  • Evaluation Frameworks:Establishrigorous internal benchmarks using both objective metrics (WER, MCD) and subjective human-in-the-loop (MOS) testing.
Qualifications
  • PhD or MSc in Computer Science with a focus on Deep Learning, Signal Processing, or Computational Linguistics.
  • Record of Research: Published work in top-tier venues (NeurIPS, ICLR, ICASSP, Interspeech).
Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Research Scientist - Speech/Audio Machine Learning
Research Scientist - Speech/Audio Machine Learning

Ifm Us • Paris

Sur place
EUR 90 000 - 130 000
Research Engineer - Speech/Audio Machine Learning
Research Engineer - Speech/Audio Machine Learning

Institute of Foundation Models • Paris

Sur place
EUR 45 000 - 75 000
Research Scientist, Speech/Audio ML — Next-Gen Audio Models
Research Scientist, Speech/Audio ML — Next-Gen Audio Models

Ifm Us • Paris

Sur place
EUR 90 000 - 130 000
Pioneering Speech & Audio ML Research Scientist
Pioneering Speech & Audio ML Research Scientist

Institute of Foundation Models • Paris

Sur place
EUR 50 000 - 70 000
Speech ML Research Engineer: Low-Latency, Scalable AI
Speech ML Research Engineer: Low-Latency, Scalable AI

Institute of Foundation Models • Paris

Sur place
EUR 45 000 - 75 000
Senior Machine Learning Scientist (ASR)
Senior Machine Learning Scientist (ASR)

Sonos, Inc. • Paris

Hybride
EUR 70 000 - 90 000
Remote AI Speech & Audio Specialist | ASR & TTS
Remote AI Speech & Audio Specialist | ASR & TTS

TELUS Digital AI Data Solutions • France

Sur place
EUR 60 000 - 90 000
VoiceAI Research Scientist - Remote/Hybrid
VoiceAI Research Scientist - Remote/Hybrid

Crane Venture Partners • Paris

Hybride
EUR 60 000 - 80 000
Competitive research compensation package
Premium health insurance
Full transportation reimbursement
+3
Research Engineer - FAIR, SGT
Research Engineer - FAIR, SGT

Meta • Paris

Sur place
EUR 90 000 - 130 000
Research Scientist - Full-Time
Research Scientist - Full-Time

Crane Venture Partners • Paris

Hybride
EUR 60 000 - 80 000
Competitive research compensation package
Premium health insurance
Full transportation reimbursement
+3