Senior Multimodal ML Engineer — Vision & Audio for Production

White Circle

Paris

Hybride

EUR 122 613 - 201 436

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Relocation package
Hybrid work from Paris
Medical insurance (France)
All the hardware, tools & services
Subscriptions for AI agents & IDEs
Team off-sites

Résumé du poste

White Circle is hiring a Multimodal ML Engineer in Paris. You will train and ship vision, audio, video, and speech models for an AI safety platform processing 100M+ API calls monthly.

You will extend models across modalities, design experiments, and build multimodal data pipelines for training-ready datasets, including synthetic data generation. You will deploy models end-to-end, define evaluation benchmarks, and optimize for production latency.

Qualifications

  • 3+ years training large-scale deep learning models in multimodal domains.
  • Strong PyTorch skills with hands-on distributed training experience (DeepSpeed, FSDP, or similar).
  • Deep experience with multimodal architectures and how encoders/projectors/LLMs fit together.
  • Hands-on with RLHF/alignment for multimodal beyond text.
  • Experience with video/audio sequence modeling and long-context processing.
  • Track record of shipping models to production with latency optimization.
  • Comfort with large-scale multimodal dataset curation and synthetic data generation.
  • Familiar with MoE architectures and their tradeoffs for multimodal workloads.
  • Strong engineering fundamentals: clean code, version control, testing, documentation

Responsabilités

  • Train and fine‑tune large‑scale multimodal models from scratch or pretrained checkpoints.
  • Extend models across modalities: vision-language, audio, video, speech.
  • Design experiments: architecture changes, data mixes, training recipes.
  • Build and maintain multimodal data pipelines from raw data to training-ready datasets.
  • Train and optimize MoE architectures for efficient inference.
  • Build alignment pipelines: SFT, DPO, reward modeling across modalities.
  • Optimize models for production: quantization, distillation, batching, streaming.
  • Deploy models end‑to‑end from research checkpoint to production serving.
  • Define and monitor evaluation metrics and benchmarks for the product

Connaissances

Deep learning
PyTorch
Distributed training
Multimodal architectures
RLHF/alignment
Video/audio sequence modeling
MoE architectures
Model deployment
Data curation
Engineering fundamentals

Outils

DeepSpeed
FSDP
PyTorch tooling

Description du poste

White Circle is hiring a Multimodal ML Engineer in Paris. You will train and ship vision, audio, video, and speech models for an AI safety platform processing 100M+ API calls monthly.

You will extend models across modalities, design experiments, and build multimodal data pipelines for training-ready datasets, including synthetic data generation. You will deploy models end-to-end, define evaluation benchmarks, and optimize for production latency.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

Multimodal ML Engineer
Multimodal ML Engineer

White Circle • Paris

Hybride
EUR 122 000 - 202 000
Relocation package
Hybrid work from Paris
Medical insurance (France)
+3
Member of Engineering (Multimodality - Research Lead)
Member of Engineering (Multimodality - Research Lead)

Jobgether • France

À distance
EUR 120 000 - 170 000
Fully remote work environment
Flexible working hours
Health insurance allowance
+4
Research Engineer — Multimodal AI & Agentic Systems
Research Engineer — Multimodal AI & Agentic Systems

H Company • Paris

Hybride
EUR 70 000 - 100 000
Competitive salary
Opportunities for professional growth
Collaborative and dynamic work environment
Senior Multimodal AI/ML Engineer - Remote
Senior Multimodal AI/ML Engineer - Remote

Overt Minds • Auvergne-Rhône-Alpes

Sur place
EUR 90 000 - 130 000
Remote work options
Professional growth
Flexible hours
+1
Lead, Multimodal AI Research (Remote)
Lead, Multimodal AI Research (Remote)

Jobgether • France

À distance
EUR 120 000 - 170 000
Fully remote work environment
Flexible working hours
Health insurance allowance
+4
Senior ML Engineer, Voice Control & Speech (Hybrid, Paris)
Senior ML Engineer, Voice Control & Speech (Hybrid, Paris)

Sonos, Inc. • Paris

Hybride
EUR 70 000 - 90 000
Enterprise AI/ML Engineer: LLMs, Production & Scale Paris
Enterprise AI/ML Engineer: LLMs, Production & Scale Paris

vector8 • Paris

Sur place
EUR 90 000 - 130 000
Flexible working hours
Hybrid model options
Private health and life insurance
+3
Senior ML Engineer - End-to-End ML in Production (Hybrid)
Senior ML Engineer - End-to-End ML in Production (Hybrid)

Voodoo • Paris

Hybride
EUR 90 000 - 130 000
Foundational ML Engineer - Multimodal SSL & Edge AI
Foundational ML Engineer - Multimodal SSL & Edge AI

Harmattan AI • Paris

Sur place
EUR 80 000 - 120 000
Senior ML Engineer, Identity & Audience — Hybrid Paris
Senior ML Engineer, Identity & Audience — Hybrid Paris

Vibe • Paris

Hybride
EUR 110 000 - 160 000
Variable pay
Hybrid work in Paris
Health insurance
+3