AI Engineer, Product

Mistral AI

Paris

Sur place

EUR 65 000 - 95 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Résumé du poste

Mistral AI in Paris is seeking a software/ML engineer to join a product team, focusing on evaluating and improving AI-powered features across search, chat, documents, and audio. You’ll define what good looks like, measure it, run experiments, and ship improvements that raise quality, latency, safety, and reliability.

You will design prompts, orchestrations, and system prompts, run A/B tests, build observability and release processes, and collaborate with Science to push production-ready AI

Qualifications

  • 3–4 years of experience in ML or software engineering with AI production exposure.
  • Strong TypeScript or Python skills.
  • Production LLM experience: prompts, tool calls, system prompts.
  • Hands-on with evals and A/B testing.
  • Observability experience: logging, tracing, dashboards, alerting.
  • Product mindset: form hypotheses, run experiments, ship.
  • Clear communication and autonomy with production impact.

Responsabilités

  • Design and run evaluations for your product area: reference tests, heuristics, model-graded checks tailored to search relevance, chat quality, document understanding, or audio performance.
  • Define and track metrics that matter: task success, helpfulness, latency, safety flags, cost.
  • Own prompt and orchestration design: write, test, and iterate on prompts and system prompts as a core part of your work.
  • Run A/B tests on prompts, models, and configurations; analyze results; make rollout or rollback decisions from data.
  • Set up observability for LLM calls: structured logging, tracing, dashboards.
  • Operate model releases: canary and shadow traffic, sign-offs, SLO-based rollback criteria.
  • Improve core behaviors in your product area: memory policies, retrieval quality, routing, tool-call reliability.
  • Create templates and documentation so other teams can author evals and ship safely.
  • Partner with Science to diagnose regressions and lead post-mortems.

Connaissances

TypeScript
Python
Production ML
Eval & A/B testing
Observability
Product mindset
Prompt engineering

Outils

Logging
Dashboards

Description du poste

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems—across high‑stakes industries like finance, manufacturing, defense, healthcare, and the public sector—co‑creating customized AI systems that they can run on their terms.

We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low‑ego and team‑spirited.

The Role

Embedded directly in a product team as search, chat, documents, or audio, you’ll improve AI‑powered features through rigorous evaluation, prompt and orchestration design, and rapid experimentation. You’ll own your domain’s AI quality end‑to‑end: define what "good" looks like, measure it, run experiments, and ship what works. Work with Science to deliver measurable improvements to quality, latency, safety, and reliability.

What You Will Do
  • Design and run evaluations for your product area: reference tests, heuristics, model‑graded checks tailored to search relevance, chat quality, document understanding, or audio performance.
  • Define and track metrics that matter: task success, helpfulness, hallucination proxies, safety flags, latency, cost.
  • Own prompt and orchestration design: write, test, and iterate on prompts and system prompts as a core part of your work.
  • Run A/B tests on prompts, models, and configurations; analyze results; make rollout or rollback decisions from data.
  • Set up observability for LLM calls: structured logging, tracing, dashboards, alerts.
  • Operate model releases: canary and shadow traffic, sign‑offs, SLO‑based rollback criteria, regression detection.
  • Improve core behaviors in your product area, whether that’s memory policies, intent classification, routing, tool‑call reliability, or retrieval quality.
  • Create templates and documentation so other teams can author evals and ship safely.
  • Partner with Science to diagnose regressions and lead post‑mortems.
What We’re Looking For
  • 3‑4 years of experience; backgrounds that fit well include ML engineers moving closer to product, or software engineers with real AI/ML production experience.
  • Strong TypeScript or Python skills – we have both tracks depending on team fit.
  • Production LLM experience: prompts, tool/function calling, system prompts.
  • Hands‑on with evals and A/B testing; you can design metrics, not just run them.
  • Comfortable implementing directly in product code, not only notebooks.
  • Observability experience: logging, tracing, dashboards, alerting.
  • Product mindset: form hypotheses, run experiments, interpret results, ship.
  • Clear communication, autonomous, and oriented toward production impact over experimentation for its own sake.
It would be ideal if you also have:
  • Safety systems experience: moderation, PII handling/redaction, guardrails.
  • Release operations: canary/shadowing, automated rollbacks, experiment platforms.
  • Prior work on search ranking, chat systems, document AI, or audio ML features.
What We Offer

We offer a comprehensive benefits package designed to support your well‑being, growth, and work‑life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location‑specific perks.

For the most up‑to‑date details on benefits available in your location, please refer to our Benefits page.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

AI Engineer, Product
AI Engineer, Product

Mistral • Paris

Sur place
EUR 70 000 - 110 000
Healthcare coverage
Relocation support
Retirement plans
+2
Applied AI Engineer, Fullstack Software Engineer - EMEA
Applied AI Engineer, Fullstack Software Engineer - EMEA

Mistral • Paris

Sur place
EUR 70 000 - 110 000
Healthcare coverage
Relocation support
Wellbeing programs
Applied AI, Forward Deployed Machine Learning Engineer - EMEA
Applied AI, Forward Deployed Machine Learning Engineer - EMEA

Mistral • Paris

Sur place
EUR 75 000 - 110 000
Engineering Manager
Engineering Manager

Mistral AI • Paris

Sur place
EUR 110 000 - 170 000
Applied Scientist/Research Engineer, EMEA
Applied Scientist/Research Engineer, EMEA

Mistral • Paris

Sur place
EUR 90 000 - 130 000
Healthcare coverage
Parental leave
Retirement plans
+3
Applied AI Engineer, Fullstack Software Engineer - EMEA
Applied AI Engineer, Fullstack Software Engineer - EMEA

Mistral AI • Paris

Sur place
EUR 55 000 - 90 000
Applied AI, Technical Lead, Forward Deployed AI Engineer
Applied AI, Technical Lead, Forward Deployed AI Engineer

Mistral • Paris

Sur place
EUR 120 000 - 180 000
Applied AI Engineer, Prototyping
Applied AI Engineer, Prototyping

Mistral • Paris

Sur place
EUR 90 000 - 120 000
Healthcare coverage
Parental leave
Relocation support
+1
AI Scientist - Agentic Engineering
AI Scientist - Agentic Engineering

Mistral • Paris

Hybride
EUR 90 000 - 120 000
Healthcare coverage
Relocation support
Wellness programs
Applied Scientist / Research Engineer, AI4Engineering
Applied Scientist / Research Engineer, AI4Engineering

Mistral • Paris

Sur place
EUR 90 000 - 150 000
Healthcare coverage
Parental leave
Retirement plans
+2