# Cité par l'IA — robots.txt # Ce site cherche à être cité par les moteurs de réponse, # pas à servir de corpus d’entraînement. # Les robots de citation sont autorisés, ceux d’entraînement refusés. User-agent: * Content-Signal: search=yes, ai-input=yes, ai-train=no Allow: / Disallow: /wp-admin/ Allow: /wp-admin/admin-ajax.php # — Moteurs de réponse et de recherche : autorisés — # Google Search User-agent: Googlebot Allow: / # Bing et Copilot User-agent: Bingbot Allow: / # OpenAI — index de ChatGPT Search User-agent: OAI-SearchBot Allow: / # OpenAI — récupération à la demande User-agent: ChatGPT-User Allow: / # Anthropic — index de recherche User-agent: Claude-SearchBot Allow: / # Anthropic — récupération à la demande User-agent: Claude-User Allow: / # Perplexity — index de recherche User-agent: PerplexityBot Allow: / # Perplexity — récupération à la demande User-agent: Perplexity-User Allow: / # Mistral — index de recherche User-agent: MistralAI-Index Allow: / # Mistral — récupération à la demande User-agent: MistralAI-User Allow: / # Amazon — recherche d’Alexa User-agent: Amzn-SearchBot Allow: / # Amazon — requêtes d’Alexa User-agent: Amzn-User Allow: / # Meta — recherche de Meta AI User-agent: meta-webindexer Allow: / # Meta — récupération à la demande User-agent: meta-externalfetcher Allow: / # DuckDuckGo — assistant User-agent: DuckAssistBot Allow: / # Apple — Siri, Spotlight et Safari User-agent: Applebot Allow: / # — Collecte pour entraînement de modèles : refusée — # Cohérent avec Content-Signal (ai-train=no) et avec la réserve # de droits TDM publiée dans /.well-known/tdmrep.json. # OpenAI — entraînement User-agent: GPTBot Disallow: / # Anthropic — entraînement User-agent: ClaudeBot Disallow: / # Google — entraînement et grounding Gemini User-agent: Google-Extended Disallow: / # Apple — entraînement d’Apple Intelligence User-agent: Applebot-Extended Disallow: / # Meta — entraînement User-agent: meta-externalagent Disallow: / # Amazon — amélioration des services, entraînement possible User-agent: Amazonbot Disallow: / # Mistral — entraînement User-agent: MistralAI-Training Disallow: / # Common Crawl — corpus d’entraînement User-agent: CCBot Disallow: / Sitemap: https://cite-par-lia.fr/sitemap.xml