Applied AI Engineer, Site Reliability Engineer - EMEA

Mistral

München

Vor Ort

EUR 90.000 - 140.000

Vollzeit

14 Tage+

Erhalte mehr Antworten von Arbeitgebern

Versende in nur wenigen Minuten einen passgenauen Lebenslauf.

Zusammenfassung

Mistral is seeking founding engineers for the Applied AI SRE sub-team in Munich. You will build and operate a framework to ensure reliable, scalable delivery across both Mistral-hosted and customer-hosted deployments.

You will own on-call, incident response, and security baselines, enabling on-demand provisioning and end-to-end reliability across multi-tenant environments. Strong Kubernetes, IaC, and programming skills are essential to succeed.

Qualifikationen

  • 5+ years in SRE, Production Engineering, or DevOps with a track record of shipping tooling.
  • Strong multi-tenant Kubernetes fluency, namespace segmentation, network policy, RBAC, admission control at scale.
  • On-call discipline: incident response, blameless post-mortems, runbook-first mindset.
  • Observability stack in production: Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz.
  • Infrastructure as code: Terraform, Ansible or equivalents.
  • Proficient in Python and/or Golang for tooling and automation.

Aufgaben

  • BUILD the fleet of Mistral platforms and apps with proactivity and reliability, create SLO templates and runbooks.
  • RUN Tier-1 customer environments, ensure SLO compliance, own on-call and incident response.
  • ENABLE productization of deployment, security baselines, and automation across Applied AI solutions.
  • SECURE oversee security operations for customer deployments, CVE response, SBOM and provenance controls.

Kenntnisse

SRE/DevOps
Kubernetes
Observability
Python/Golang
On-call / Incident
Security mindset

Tools

Terraform
Ansible
OpenTelemetry
Loki
Tempo
Signoz

Jobbeschreibung

About Mistral

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems—across high-stakes industries like finance, manufacturing, defense, healthcare, and the public sector—co-creating customized AI systems that they can run on their terms.

About The Team

The Applied AI team is Mistral's customer-facing technical organization. We work directly with enterprise clients from pre-sales through implementation to deploy cutting-edge AI solutions that deliver measurable business impact.

Our team combines deep ML expertise with strong customer engagement skills, operating like startup CTOs who own end-to-end project execution. Our SRE team works transversally across customer engagements, enabling value creation through Mistral tech - at scale. By joining the team you will bridge the gap between cutting-edge AI research and real-world enterprise applications, ensuring our solutions are robust, scalable, and aligned with both customer needs and Mistral's technological vision.

About The Job

You will be one of the founding engineers of the Applied AI SRE sub-team.

Your mission, alongside the team, is to build and operate the framework to ensure Mistral’s solution delivery is reliable and sustainable - and applied uniformly across all our accounts, both Mistral-hosted and customer-hosted. You should already have a strong understanding of what operational excellence looks like, and you’re ready to scale your impact.

  • BUILD – Design for a fleet of Mistral platforms and apps. Build proactivity to reduce reactivity. Productize reliability, author runbooks, create SLO templates, implement observability.
  • RUN – Operate the Tier-1 customer environments that Mistral are contracted to operate. Ensure SLO compliance, own on-call and incident response, manage drift, partner with Technical Support as L3 escalation, champion high signal post-mortems.
  • ENABLE – Productize how Mistral deploy, secure, and scale our Applied AI solutions. Engineer on-demand provisioning, author security baseline packages, embed security guardrails, automate everything.
  • SECURE – Own the security operations layer for our customer-side deployments. Lead CVE response across the fleet, ship supply-chain integrity controls (SBOM, signed images, provenance), co-page with InfoSec on security incidents, enforce secure-config baselines.

This is a framework-first, fleet management role at heart. If you're excited by the difference between solving one customer's problem and structurally solving the class of problem for every customer, this is the role.

How We Work In Applied AI
  • We care about people and outputs.
  • What matters is what you ship, not the time you spend on it.
  • Bureaucracy is where urgency goes to vanish. You talk to whoever you need to talk to. The best idea wins, whether it comes from a principal engineer or someone in their first week.
  • Always ask why. The best solutions come from deep understanding, not from copying what worked before.
  • We say what we mean. Feedback is direct, timely, and given because we care.
  • No politics. Low ego, high standards.
  • We embrace an unstructured environment and find joy in it.
About You
  • Fluent in English.
  • 5+ years in SRE, Production Engineering, or DevOps, with a record of shipping tooling.
  • Strong multi-tenant Kubernetes fluency, namespace segmentation, network policy, RBAC, admission control, operations at scale.
  • On-call discipline: incident response, blameless post-mortem culture, runbook-first mindset.
  • Observability stack in production: Prometheus, Grafana, OpenTelemetry, Loki, Tempo, Signoz.
  • Infrastructure as code: Terraform, Ansible (or close equivalents).
  • Proficient in Python and/or Golang for tooling and automation.
  • Security mindset: you treat secure-SDLC, CVE response, and supply-chain integrity as reliability properties of the shipped artifact, not as someone else's job.
  • Strong written communication skills: runbooks, post-mortems, and customer-facing incident comms are core deliverables of this role.
  • Comfortable operating with high autonomy in an ambiguous, fast-paced environment — and disciplined enough to defend the team's scope when work tries to spill in.
  • Solid Linux internals, networking debug, and distributed-systems fundamentals.
Strong plus
  • Cloud or application security background (AppSec, K8s security, supply chain — SBOM, cosign, SLSA). At least one of our early hires must bring this; if it's you, flag it.
  • Experience operating LLM / model-serving stacks in production
  • Experience with multi-cloud or on-prem hybrid customer environments (AWS, GCP, Azure, sovereign clouds).
  • Open-source contributions, particularly in SRE, observability, or security tooling.
What we offer

We offer a comprehensive benefits package designed to support your well-being, growth, and work-life balance. Benefits vary by country and may include healthcare coverage, parental leave, retirement plans, relocation support, wellness programs, meal and transportation allowances, and other location-specific perks. For the most up-to-date details on benefits available in your location, please refer to our Benefits page.

Privacy Policy

Your privacy matters to us. You can learn more about how we handle your personal data in our Applicant Privacy Policy. By applying, you agree to our Applicant Privacy Policy.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.
oder ziehe deine Datei hierhin.
Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Applied AI Engineer, Site Reliability Engineer - EMEA
Applied AI Engineer, Site Reliability Engineer - EMEA

Mistral AI • München

Vor Ort
EUR 90.000 - 140.000
Applied AI, Forward Deployed Machine Learning Engineer
Applied AI, Forward Deployed Machine Learning Engineer

Mistral Ai • München

Vor Ort
EUR 70.000 - 110.000
Site Reliability Engineer, Mistral Cloud
Site Reliability Engineer, Mistral Cloud

Mistral • München

Vor Ort
EUR 90.000 - 130.000
Site Reliability Engineer, Mistral Cloud
Site Reliability Engineer, Mistral Cloud

Mistral • Berlin

Vor Ort
EUR 70.000 - 120.000
Applied AI, Forward Deployed Machine Learning Engineer
Applied AI, Forward Deployed Machine Learning Engineer

Mistral • München

Vor Ort
EUR 90.000 - 130.000
Applied AI, Technical Lead, Forward Deployed AI Engineer
Applied AI, Technical Lead, Forward Deployed AI Engineer

Mistral AI • München

Vor Ort
EUR 120.000 - 180.000
AI Deployment Strategist
AI Deployment Strategist

Mistral Ai • München

Vor Ort
EUR 90.000 - 130.000
Healthcare coverage
Parental leave
Relocation support
+1
Applied AI, Technical Lead, Forward Deployed AI Engineer
Applied AI, Technical Lead, Forward Deployed AI Engineer

Mistral • München

Vor Ort
EUR 120.000 - 180.000
AI Deployment Strategist, Physics - EMEA
AI Deployment Strategist, Physics - EMEA

Mistral • München

Vor Ort
EUR 90.000 - 140.000
AI Deployment Strategist
AI Deployment Strategist

Mistral • München

Vor Ort
EUR 90.000 - 140.000
Healthcare coverage
Relocation support
Wellness programs