Visits a web page when a user asks Mistral's assistant a question, and may link the source in the answer. Mistral says it is not used for automatic crawling or for training.
Source: docs.mistral.aiOperator
Mistral
Job
User-triggered fetch
Visits a page because a person asked a question that needs it. This is the visit that can end in a citation.
robots.txt
Follows robots.txt
Mistral publishes a robots.txt token for each of its crawlers and documents robots.txt as the way to control them.
Source: docs.mistral.aiMozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; MistralAI-User/1.0; +https://docs.mistral.ai/robots)
Add one of these groups to the robots.txt at the root of your domain.
A crawler obeys only the most specific group that names it (RFC 9309), so a named group does not inherit the rules under User-agent: *. Repeat any paths you keep private, as the example does with /admin/.
User-agent: MistralAI-User Allow: / Disallow: /admin/
User-agent: MistralAI-User Disallow: /
Read from the robots.txt we serve. We want AI engines to read and cite this site, so every AI crawler is allowed everywhere except the app, the API and the sign-in pages (/dashboard/, /api/, /login, /signup, /sign-in, /sign-up). The file also declares Content-Signal: ai-train=yes, search=yes, ai-input=yes.
User-agent: MistralAI-User Allow: / Disallow: /dashboard/ Disallow: /api/ Disallow: /login Disallow: /signup Disallow: /sign-in Disallow: /sign-up
See the whole file. The same rules repeat in every group, because a crawler that matches a named group ignores the wildcard one.
Sources, checked 2026-09-24