Research Engineer, Forge

Mistral AI

Greater London

Vor Ort

GBP 110.000 - 170.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Hebe dich für diese Rolle von der Masse ab — erstelle in etwa einer Minute einen maßgeschneiderten Lebenslauf und ein Anschreiben.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Healthcare coverage
Parental leave
Relocation support
Wellness programs
Meal and transportation allowances

Zusammenfassung

Mistral AI in London seeks a Research Engineer for the Forge team to translate customer requirements into reliable training and deployment workflows across CPT/SFT/RL/distillation, evaluation, data, and infrastructure.

This role sits in Applied Science and collaborates with scientists, engineers, product, and customer-facing teams to ensure Forge projects ship, are maintainable, and trustworthy. Interview focus varies; strong Python, PyTorch/JAX, and infra fundamentals are valuable.

Qualifikationen

  • Strong Python engineering skills and experience working in large codebases (testing, code review, CI, operational ownership).
  • Hands-on experience with PyTorch, JAX, or similar.
  • Strong systems and infrastructure fundamentals.
  • Experience with LLM training or post-training: fine-tuning, RL, distillation, evaluation, and/or data pipelines.
  • Excellent debugging skills in ambiguous systems (distributed jobs, data issues, quality regressions, infra failures).
  • Clear communication with technical and non-technical stakeholders.
  • High agency, low ego, and comfort in fast-moving, under-specified environments.

Aufgaben

  • Build and improve post-training and evaluation workflows (CPT/SFT/RL/distillation), turning prototypes into repeatable Forge “recipes”.
  • Develop tools and pipelines for synthetic data generation, data curation, training, evaluation, and deployment.
  • Debug and harden large-scale ML systems: distributed training, scheduling/execution, checkpointing, observability, and reproducibility.
  • Improve the Forge codebase via clear APIs, tests, documentation, and maintainable abstractions.
  • Push the frontier of our RL training stack (e.g., high-throughput async rollout and scalable post-training systems at frontier-model scale)
  • Make sure Forge deployment is seamless and adaptable to a diversity of clients (hardware access, software stack, cloud and on-premises, …)
  • Partner with researchers and infrastructure engineers to translate bottlenecks into concrete system improvements.

Kenntnisse

Strong Python engineering skills
PyTorch or JAX
Systems and infrastructure foundations
LLM training / post-training (fine-tun
Debugging distributed systems
Clear communication with stakeholders
High agency, low ego

Jobbeschreibung

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector, co-creating customized AI systems that they can run on their terms.

We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Our teams are distributed between Europe, North America, Asia and the Middle East. We are creative, low-ego and team-spirited.

Role summary

As a Research Engineer on Forge, you will turn real customer requirements into reliable training and deployment workflows. You’ll work end‑to‑end across model adaptation and post‑training (CPT/SFT/RL/distillation), evaluation, data, and infrastructure. The role bridges research experimentation and production constraints.

This role sits in Applied Science, with direct impact on client outcomes. You’ll collaborate closely with scientists, engineers, product, and customer‑facing teams to ensure Forge projects ship, are maintainable, and can be trusted by others.

Interview focus can vary (algorithms, infrastructure, evals, or data). You don’t need to match every bullet below to apply.

What you will do
  • Build and improve post‑training and evaluation workflows (CPT/SFT/RL/distillation), turning prototypes into repeatable Forge “recipes”.

  • Develop tools and pipelines for synthetic data generation, data curation, training, evaluation, and deployment.

  • Debug and harden large‑scale ML systems: distributed training, scheduling/execution, checkpointing, observability, and reproducibility.

  • Improve the Forge codebase via clear APIs, tests, documentation, and maintainable abstractions.

  • Push the frontier of our RL training stack (e.g., high-throughput async rollout and scalable post‑training systems at frontier-model scale)

  • Make sure Forge deployment is seamless and adaptable to a diversity of clients (hardware access, software stack, cloud and on-premises, …)

  • Partner with researchers and infrastructure engineers to translate bottlenecks into concrete system improvements.

About you
  • Strong Python engineering skills and experience working in large codebases (testing, code review, CI, operational ownership).

  • Hands‑on experience with PyTorch, JAX, or similar.

  • Strong systems and infrastructure fundamentals.

  • Experience with LLM training or post‑training: fine‑tuning, RL, distillation, evaluation, and/or data pipelines.

  • Excellent debugging skills in ambiguous systems (distributed jobs, data issues, quality regressions, infra failures).

  • Clear communication with technical and non‑technical stakeholders.

  • High agency, low ego, and comfort in fast‑moving, under‑specified environments.

Nice to have
  • Distributed training experience (FSDP, DeepSpeed, Megatron, etc.).

  • Cluster/orchestration experience (SLURM, Ray, Kubernetes, Kueue, Karpenter, Skypilot, etc.).

  • Experience building reliable ML infrastructure, evaluation systems, or large‑scale data processing pipelines.

  • Research experience in LLMs, agents, multimodal models, reasoning, code, or domain adaptation.

  • Open‑source contributions, publications, or widely used internal tooling.

  • Experience training multi‑billion‑parameter models (pre‑training or RL). Experience training on petabyte- and exabyte-scale datasets.

  • Ability to identify bottlenecks across the stack and drive improvements from first principles.

What We Offer

We offer a comprehensive benefits package designed to support your well‑being, growth, and work‑life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location‑specific perks.

For the most up‑to‑date details on benefits available in your location, please refer to our Benefits page.

Privacy Policy

Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Research Engineer, Forge
Research Engineer, Forge

Mistral • Greater London

Vor Ort
GBP 90.000 - 140.000
Research Engineer, Machine Learning
Research Engineer, Machine Learning

Mistral AI • Greater London

Vor Ort
GBP 120.000 - 180.000
Healthcare coverage
Relocation support
Wellness programs
Research Software Engineer
Research Software Engineer

Mistral • Greater London

Vor Ort
GBP 90.000 - 120.000
Research Engineer, Machine Learning
Research Engineer, Machine Learning

Mistral • Greater London

Vor Ort
GBP 100.000 - 140.000
Healthcare coverage
Parental leave
Retirement plans
+3
Research Engineer - AI Systems & Production Pipelines
Research Engineer - AI Systems & Production Pipelines

Mistral AI • Greater London

Vor Ort
GBP 110.000 - 170.000
Healthcare coverage
Parental leave
Relocation support
+2
Applied Scientist, EMEA
Applied Scientist, EMEA

Mistral AI • Greater London

Vor Ort
GBP 100.000 - 170.000
Healthcare coverage
Parental leave
Relocation support
+3
AI Scientist
AI Scientist

Mistral AI • Greater London

Vor Ort
GBP 110.000 - 170.000
Healthcare coverage
Relocation support
Wellness programs
AI Scientist - Agentic Engineering
AI Scientist - Agentic Engineering

Mistral AI • Greater London

Vor Ort
GBP 120.000 - 180.000
Healthcare coverage
Relocation support
Wellness programs
+1
Research Engineer, Inference Foundation
Research Engineer, Inference Foundation

Mistral • Greater London

Hybrid
GBP 120.000 - 160.000
Research Engineer, Data Infrastructure
Research Engineer, Data Infrastructure

Mistral • Greater London

Vor Ort
GBP 76.673 - 110.750
Healthcare coverage
Relocation support
Wellness programs