Research Engineer - Speech/Audio Machine Learning

Ifm Us

Paris

Sur place

EUR 70 000 - 95 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

Institute of Foundation Models seeks a Research Engineer specializing in speech/audio machine learning to bridge research and production, delivering scalable, low-latency AI systems.

You will optimize compute graphs, apply model compression, build reliable data pipelines for audio datasets, and implement distributed training using DeepSpeed and Horovod, with deployment in Docker and WebRTC integrations.

Qualifications

  • MSc or BSc in Computer Science or Software Engineering.
  • Engineering Excellence: Strong background in Systems Programming and Computer Architecture.

Responsabilités

  • Compute Optimization: Optimize model compute graphs for target runtimes using frameworks such as TensorRT, Apache TVM, or OpenVINO.
  • Model Compression: Perform quantization (INT8/FP8) and pruning for real-time execution on edge or cloud GPUs.
  • Data Orchestration: Build reliable data pipelines for audio datasets, including automated denoising and synthetic data augmentation.
  • Distributed Training: Help design efficient multi-node training scripts using frameworks such as DeepSpeed, Horovod, as well as other distributed training technologies.
  • Deployment: Build and maintain high-performance demo environments on the cloud using Docker and standalone model serving frameworks. Integrate WebRTC for real-time audio-to-audio interactions.

Connaissances

Compute optimization
Model compression
Data orchestration
Distributed training
Deployment

Formation

MSc or BSc in Computer Science or Software Engineering

Outils

TensorRT
Apache TVM
OpenVINO
DeepSpeed
Horovod
Docker
WebRTC

Description du poste

About the Institute of Foundation Models

We are a dedicated research lab for building, understanding, using, and risk-managing foundation models. Our mandate is to advance research, nurture the next generation of AI builders, and drive transformative contributions to a knowledge-driven economy.

As part of our team, you’ll have the opportunity to work on the core of cutting-edge foundation model training, alongside world-class researchers, data scientists, and engineers, tackling the most fundamental and impactful challenges in AI development. You will participate in the development of groundbreaking AI solutions that have the potential to reshape entire industries. Strategic and innovative problem-solving skills will be instrumental in establishing MBZUAI as a global hub for high-performance computing in deep learning, driving impactful discoveries that inspire the next generation of AI pioneers.

The Role

As a Research Engineer specializing in speech/audio machine learning, you will bridge the gap between research and production constraints to deliver competitive performance. You will be responsible for the end-to-end performance of our machine learning speech systems, from architecting scalable training infrastructure to achieving extreme low-latency inference.

Responsibilities
  • Compute Optimization: Optimize model compute graphs for target runtimes using frameworks such as TensorRT, Apache TVM, or OpenVINO.
  • Model Compression: Perform quantization (INT8/FP8) and pruning to enable real-time execution on edge or cloud GPUs.
  • Data Orchestration: Build reliable data pipelines for audio datasets, including automated denoising and synthetic data augmentation.
  • Distributed Training: Help design efficient multi-node training scripts using frameworks such as DeepSpeed, Horovod, as well as other distributed training technologies.
  • Deployment: Build and maintain high-performance demo environments on the cloud using Docker and standalone model serving frameworks. Integrate WebRTC or similar bidirectional protocols to ensure seamless, real-time audio-to-audio interactions.
  • MSc or BSc in Computer Science or Software Engineering.
  • Engineering Excellence: Strong background in Systems Programming and Computer Architecture.
Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Research Engineer - Speech/Audio Machine Learning
Research Engineer - Speech/Audio Machine Learning

Institute of Foundation Models • Paris

Sur place
EUR 45 000 - 75 000
Research Scientist - Speech/Audio Machine Learning
Research Scientist - Speech/Audio Machine Learning

Ifm Us • Paris

Sur place
EUR 90 000 - 130 000
Research Scientist - Speech/Audio Machine Learning
Research Scientist - Speech/Audio Machine Learning

Institute of Foundation Models • Paris

Sur place
EUR 50 000 - 70 000
Speech & Audio ML Research Engineer — Real-Time Edge Deployment
Speech & Audio ML Research Engineer — Real-Time Edge Deployment

Ifm Us • Paris

Sur place
EUR 70 000 - 95 000
Research Scientist, Speech/Audio ML — Next-Gen Audio Models
Research Scientist, Speech/Audio ML — Next-Gen Audio Models

Ifm Us • Paris

Sur place
EUR 90 000 - 130 000
Speech ML Research Engineer: Low-Latency, Scalable AI
Speech ML Research Engineer: Low-Latency, Scalable AI

Institute of Foundation Models • Paris

Sur place
EUR 45 000 - 75 000
AI Scientist - Audio
AI Scientist - Audio

Mistral • Paris

Sur place
EUR 55 000 - 75 000
Competitive cash salary and equity
Daily lunch vouchers
Monthly contribution to gym membership
+3
Pioneering Speech & Audio ML Research Scientist
Pioneering Speech & Audio ML Research Scientist

Institute of Foundation Models • Paris

Sur place
EUR 50 000 - 70 000
Research Engineer - FAIR, SGT
Research Engineer - FAIR, SGT

Meta • Paris

Sur place
EUR 90 000 - 130 000
Software Engineer, Data Infrastructure & Acquisition - Lille, France
Software Engineer, Data Infrastructure & Acquisition - Lille, France

Clutch Canada • Lille

Sur place
EUR 68 000 - 103 000
Competitive salaries
Laid-back atmosphere
Opportunity to impact learning differences