Research Scientist - Full-Time

Crane Venture Partners

Paris

Sur place

EUR 60 000 - 80 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Competitive research compensation package
Premium health insurance
Full transportation reimbursement
5 weeks of paid vacation + 10 RTT days
Conference travel budget
Full remote flexibility

Résumé du poste

A leading research firm is seeking a Research Scientist (VoiceAI) to expand its pioneering work in speaker intelligence. The candidate will conduct cutting-edge research on multi-talker speech processing and publish findings in prestigious conferences. Ideal applicants have a PhD in relevant fields and experience in training models using PyTorch. This role offers immediate impact, competitive compensation, and unique resources including access to advanced supercomputing facilities. Join a collaborative team and shape the future of voice AI.

Qualifications

  • PhD in a relevant field is required.
  • Experience in training neural networks is necessary.
  • Publication track record in top-tier conferences required.

Responsabilités

  • Conduct research in multi-talker conversational speech processing.
  • Train state-of-the-art models using PyTorch.
  • Publish research findings in top-tier conferences.

Connaissances

PhD in speech processing
Experience training models in PyTorch
Strong software engineering fundamentals
Track record of publication at conferences
Collaborative development experience
Fluent in English

Formation

PhD in speech processing, audio understanding, or machine learning

Outils

PyTorch
Git
Python

Description du poste

Role: Research Scientist (VoiceAI)

Location: Toulouse, Paris, Remote
Job type: Full-time
Work setup: On-Site, Hybrid, Remote
Start: ASAP

Job offer
About pyannoteAI

pyannoteAI is pioneering Speaker Intelligence AI, transforming how AI processes and understands spoken language. Our speaker diarization technology distinguishes speakers with unmatched precision, regardless of the spoken language, making AI understand not just what is said, but who said it and when.

Founded by voice AI experts with 10+ years in the industry (ex-CNRS research scientists), we\'ve built the 9th most downloaded open-source model on HuggingFace with 52 million monthly downloads and over 140,000 users worldwide. After raising €8M from leading international VCs (Crane Venture Partners, Serena, and angels from HuggingFace and OpenAI), we\'re now scaling our enterprise platform.

From meeting transcription and call center analytics to video dubbing and voice agents, pyannoteAI powers the next generation of voice-enabled applications across industries that depend on understanding who speaks and when.

Your role

As a Research Scientist at pyannoteAI, you\'ll push the boundaries of multi-talker conversational speech processing, conducting both applied and moonshot research that gets published at top-tier conferences and deployed to production. Working alongside the creators of pyannote.audio, you\'ll train state-of-the-art models at scale and collaborate with engineering to bring breakthrough research to 140K+ developers and large Enterprise customers worldwide.

You\'ll:

  • Conduct cutting-edge research - Explore batch and streaming approaches to multi-talker speech processing including speaker diarization, speaker separation, multi-talker transcription, speaker identification, and speaker profiling
  • Train models at scale - Design, implement, and train neural network architectures using PyTorch
  • Publish at top-tier venues - Contribute original research to leading speech and machine learning conferences
  • Bridge research and production - Collaborate with our tech team to optimize and deploy models that serve millions of API calls
  • Contribute to our codebase - Develop and maintain both our proprietary and open-source libraries

What makes this role unique: Unlike typical research roles, you\'ll see your work both published AND deployed to production. With access to Jean Zay supercomputer and mentorship from researchers who pioneered speaker diarization for over a decade, you\'ll have the resources to do your best work with immediate real-world impact.

What we’re looking for
Must-haves:
  • PhD in speech processing, audio understanding, or machine learning - Deep theoretical foundation in the core domains relevant to our work
  • Track record of publication at top-tier conferences - You\'ve published at venues like Interspeech, ICASSP, NeurIPS, ICML, or similar speech/ML conferences
  • Experience training state-of-the-art models in PyTorch - You\'ve designed and trained neural network architectures that achieve competitive or SOTA results
  • Strong software engineering fundamentals - Advanced Python, Git, PyTorch, and Jupyter notebooks with clean, reproducible code practices
  • Collaborative development experience - Comfortable with Git workflows, pull requests, code reviews, and working in a team environment
  • Fluent in English - Able to write papers, present research, and communicate with the team professionally
Nice-to-haves:
  • Experience with pyannote.audio - You\'ve used our toolkit in past research projects and understand its architecture
  • Model optimization expertise - Experience with distillation, quantization, pruning, or other inference optimization techniques
  • HPC and cluster computing - Familiarity with distributed training and large-scale computing environments, ideally slurm job scheduling
  • Rust programming - Ability to contribute to performance-critical inference systems
  • Fluent in French - Being based in France, it\'s always a plus
What you’ll get
  • Benefits: Competitive research compensation package with attractive BSPCE (French ESOP)
  • Premium Alan health insurance
  • Full transportation reimbursement
  • 5 weeks of paid vacation + 10 RTT days
  • Conference travel budget
  • Work Environment: Brand-new, premium offices in Toulouse or at La Maison in central Paris (Motier Ventures\' hub) - inspiring spaces designed for fast-growing startups
  • Full remote flexibility - Work from anywhere, with the option to join us in our Toulouse or Paris offices if you prefer in-person collaboration
  • Top-tier equipment - Latest MacBooks, GPU access for rapid prototyping and testing, and everything you need to do your best work
  • Access to Jean Zay supercomputer - Train models on one of Europe\'s most powerful AI supercomputers (managed by GENCI)
  • Research & Impact: Work with pioneering researchers - Collaborate directly with ex-CNRS scientists who created pyannote.audio and are recognized globally as leaders in speaker diarization
  • Publish and deploy - Unlike pure research roles, see your work both published at top venues AND deployed to production serving 140K+ developers
  • Immediate real-world impact - Your research improvements directly accelerate model development and unlock new capabilities used by millions of end users
  • Cutting-edge challenges - Work on diverse voice AI problems spanning multiple languages, real-time streaming, and production-scale datasets
  • Research autonomy - Freedom to pursue both applied research aligned with product needs and moonshot explorations that could redefine the field
  • Mentorship and growth - Learn from researchers who have shaped the speaker diarization field for over a decade, with potential to grow into senior research leadership
⭐️ Hiring process

We\'ve designed a comprehensive process to ensure mutual fit for this strategic role. Here\'s what to expect:

  • Screening call (30-45 min) - Get to know each other with our Chief of Staff and explore product philosophy alignment
  • Take-home research study (2-3 days) - Take home a research study focused on testing your skills and ability to think thoroughly on a research assignment
  • Research presentation (60 min) - Present the result of the research study (30min) as well as a past research of your choice (30min) to our CSO and a senior researcher of our team
  • Founders conversation (45-60 min) - Meet our CEO and and CTO co-founders to align on vision, discuss the voice AI landscape, and explore what excites you about the space
  • Timeline: Typically 2.5-3 weeks from application to offer. We\'ll keep you informed at every stage and respond within 2-3 days after each step.

Apply here: Typeform Research Scientist @pyannoteAI

Equal Opportunity Employer

pyannoteAI is committed to creating a diverse and inclusive workplace. We are an equal opportunity employer and welcome applications from all qualified candidates regardless of gender, gender identity or expression, sexual orientation, race, ethnicity, national origin, age, disability, religion, or any other characteristic protected by law.

All employment decisions at pyannoteAI are based on business needs, job requirements, and individual qualifications. We believe that diverse perspectives strengthen our team and drive innovation in voice AI technology.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Founding Machine Learning Scientist
Founding Machine Learning Scientist

Noïa Labs • Paris

Sur place
EUR 90 000 - 130 000
Equity package
Full health coverage
Public transport reimbursement (Paris)
+3
Software Engineer, Data Infrastructure & Acquisition - Paris, France
Software Engineer, Data Infrastructure & Acquisition - Paris, France

Clutch Canada • Paris

Sur place
EUR 50 000 - 80 000
Competitive salaries
Flexible work environment
Friendly atmosphere
+1
Software Engineer, Data Infrastructure & Acquisition - Bordeaux, France
Software Engineer, Data Infrastructure & Acquisition - Bordeaux, France

Clutch Canada • Bordeaux

Sur place
EUR 40 000 - 70 000
Friendly atmosphere
Competitive salaries
Flexible work culture
Software Engineer, Data Infrastructure & Acquisition - Lille, France
Software Engineer, Data Infrastructure & Acquisition - Lille, France

Clutch Canada • Lille

Sur place
EUR 68 000 - 103 000
Competitive salaries
Laid-back atmosphere
Opportunity to impact learning differences
AI Instructional Designer (Forward-Deployed, founding team)
AI Instructional Designer (Forward-Deployed, founding team)

Raison • Paris

Hybride
EUR 90 000 - 120 000
Equity stake
Remote-friendly collaboration
Office in Montmartre
Research Scientist (AI Behaviours)
Research Scientist (AI Behaviours)

Aisafety • Paris

Hybride
EUR 70 000 - 110 000
Paid time off
Relocation package
Medical insurance (France)
+3
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Sonos LLC • Paris

Hybride
EUR 80 000 - 100 000
Senior ML Manager - Voice Teams (x/f/m)
Senior ML Manager - Voice Teams (x/f/m)

Doctolib • Paris

Hybride
EUR 120 000 - 170 000
Full remote flexibility
Hybrid work model
Startup Partnerships - France & Southern Europe Anthropic Paris, France
Startup Partnerships - France & Southern Europe Anthropic Paris, France

Neura Market • Paris

Hybride
EUR 120 000 - 180 000
Senior Software Engineer, Core Experiences - Paris, France
Senior Software Engineer, Core Experiences - Paris, France

Speechify • Paris

Hybride
EUR 26 000 - 86 000
Competitive salary
Async culture
Impactful product
+1