Get more replies from employers
Send a job-specific resume in minutes.
Mistral AI is hiring for a Data Infrastructure Engineer to help architect and operate the next generation of data infrastructure for large-scale AI workloads. You will design and scale compute fleets and storage systems for high performance and secure data access across ML research and production workloads.
You will own end-to-end lifecycle—from migrating away from legacy schedulers to building production pipelines and supporting on-call rotations for critical training jobs, across Kubernetes
Mistral provides full‑stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems—across high‑stakes industries like finance, manufacturing, defense, healthcare, and the public sector—co‑creating customized AI systems that they can run on their terms.
The Data Infrastructure team at Mistral AI is architecting the backbone of our frontier model training and fine‑tuning ecosystem. We are building the specialized compute and data fabrics required to power the development of world‑class AI. Our vision is to operate some of the largest compute fleets in production and build data lakes and metadata systems with a roadmap toward exabyte‑scale architecture. We are currently building a high‑performance training platform designed for massive scale across both on‑premise and cloud‑native Kubernetes environments. We are leading a strategic transition from legacy scheduling to modern orchestration, implementing sophisticated multi‑cluster orchestration and cloud‑bursting capabilities to better utilize our global resources and ensure our researchers have seamless access to compute wherever it resides. Our mission is to evolve our current systems into a durable yet flexible platform.
Location: Paris / Warsaw / Zurich / London (hybrid) or remote EU/UK with one hub visit per month.
This role focuses on building and operating the next generation of data infrastructure at Mistral AI. You will be a core contributor to our evolution, helping us design and scale massive compute fleets and storage systems designed for high performance and scalability. You will help us move toward a future of decoupled control and data planes, scaling big data compute and storage platforms while ensuring secure and governed data access for MLOps and research. You will take full lifecycle ownership: from architecting the migration away from legacy orchestrators to implementing production‑grade pipelines and participating in on‑call rotations for critical training jobs.
This role is primarily based at one of our European offices (Paris, London, Warsaw, and Zurich). We will prioritize candidates who either reside there or are open to relocating. We strongly believe in the value of in‑person collaboration to foster strong relationships and seamless communication within our team. In certain specific situations, we will also consider remote candidates based in one of the countries listed in this job posting—currently France, UK, Poland, and Switzerland. In that case, we ask all new hires to visit the local hub: for the first week of their onboarding (accommodation and travelling covered) and then at least 3 days per month.
We offer a comprehensive benefits package designed to support your well‑being, growth, and work‑life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location‑specific perks. For the most up‑to‑date details on benefits available in your location, please refer to our Benefits page.