# Original User-agent: * Allow: / Sitemap: https://www.vdid.de/sitemap.xml # robots.txt – Blockiert KI-Crawler (Stand Mitte 2026) # Hinweis: robots.txt wirkt nur bei kooperativen Bots. # Bots, die sich nicht daran halten (z. B. Bytespider), # müssen per .htaccess blockiert werden. # ── KI-Training-Crawler ───────────────────────────── # OpenAI (Training) User-agent: GPTBot Disallow: / # Anthropic (Training) User-agent: ClaudeBot Disallow: / # Common Crawl (Datensatz, wird von vielen KI-Firmen genutzt) User-agent: CCBot Disallow: / # Google KI-Training (Gemini) – beeinflusst NICHT das normale Google-Ranking User-agent: Google-Extended Disallow: / # ByteDance / TikTok User-agent: Bytespider Disallow: / # Meta (Training) User-agent: meta-externalagent Disallow: / User-agent: FacebookBot Disallow: / # Apple KI-Training – beeinflusst NICHT die normale Applebot-Suche User-agent: Applebot-Extended Disallow: / # Amazon User-agent: Amazonbot Disallow: / # Allen Institute for AI User-agent: AI2Bot Disallow: / # Diffbot User-agent: Diffbot Disallow: / # Weitere Trainings-/Daten-Crawler User-agent: Timpibot Disallow: / User-agent: omgili Disallow: / User-agent: cohere-ai Disallow: / # ── KI-Such-Bots (Antworten mit Quellenangabe) ────── # Diese Bots holen Inhalte, wenn Nutzer in ChatGPT, Claude # oder Perplexity suchen – Ihre Seite wird dabei als Quelle # verlinkt und kann Besucher bringen. Standardmäßig hier # ERLAUBT. Zum Blockieren die Kommentarzeichen entfernen. # User-agent: OAI-SearchBot # Disallow: / # User-agent: ChatGPT-User # Disallow: / # User-agent: Claude-SearchBot # Disallow: / # User-agent: Claude-User # Disallow: / # User-agent: PerplexityBot # Disallow: / # User-agent: Perplexity-User # Disallow: / # User-agent: DuckAssistBot # Disallow: /