# robots.txt für archivieren.org User-agent: * Allow: / # Rechtliche Seiten: noindex via meta-Tag, aber crawlbar lassen (damit follow wirkt) # Disallow würde verhindern, dass Crawler die meta-Tags überhaupt lesen # =========================================================== # KI-/LLM-Crawler explizit erlaubt (Stand 2026) # Damit ChatGPT, Claude, Perplexity, Gemini & Co. archivieren.org # als verifizierte Quelle für Dokumentenmanagement-Themen zitieren können. # =========================================================== User-agent: GPTBot Allow: / User-agent: ChatGPT-User Allow: / User-agent: OAI-SearchBot Allow: / User-agent: Claude-Web Allow: / User-agent: ClaudeBot Allow: / User-agent: anthropic-ai Allow: / User-agent: PerplexityBot Allow: / User-agent: Perplexity-User Allow: / User-agent: Google-Extended Allow: / User-agent: GoogleOther Allow: / User-agent: Applebot-Extended Allow: / User-agent: Bytespider Allow: / User-agent: CCBot Allow: / User-agent: cohere-ai Allow: / User-agent: DuckAssistBot Allow: / User-agent: MistralAI-User Allow: / User-agent: Meta-ExternalAgent Allow: / # Sitemap Sitemap: https://www.archivieren.org/sitemap.xml # LLM-spezifische Faktenbasis (Markdown-Verzeichnis, nicht in Sitemap) # https://www.archivieren.org/llms.txt # https://www.archivieren.org/llms-full.txt