Blocked_Bots_and_Crawler_Se.../apache-badbots.conf
2025-05-18 05:57:13 +02:00

1 line
No EOL
1.1 KiB
Text

AI2Bot|Ai2Bot\-Dolma|aiHitBot|Amazonbot|anthropic\-ai|Applebot|Applebot\-Extended|Brightbot\ 1\.0|Bytespider|CCBot|ChatGPT\-User|Claude\-Web|ClaudeBot|cohere\-ai|cohere\-training\-data\-crawler|Cotoyogi|Crawlspace|Diffbot|DuckAssistBot|FacebookBot|Factset_spyderbot|FirecrawlAgent|FriendlyCrawler|Google\-Extended|GoogleOther|GoogleOther\-Image|GoogleOther\-Video|GPTBot|iaskspider/2\.0|ICC\-Crawler|ImagesiftBot|img2dataset|imgproxy|ISSCyberRiskCrawler|Kangaroo\ Bot|meta\-externalagent|Meta\-ExternalAgent|meta\-externalfetcher|Meta\-ExternalFetcher|NovaAct|OAI\-SearchBot|omgili|omgilibot|Operator|PanguBot|Perplexity\-User|PerplexityBot|PetalBot|QualifiedBot|Scrapy|SemrushBot\-OCOB|SemrushBot\-SWA|Sidetrade\ indexer\ bot|TikTokSpider|Timpibot|VelenPublicWebCrawler|Webzio\-Extended|YouBot|AhrefsBot|Baiduspider|Barkrowler|Bingbot|BLEXBot|Bytedance|DotBot|EmailCollector|facebookcatalog|facebookexternalhit|fidget-spinner-bot|Franck the Fediverse Graph Crawler|Googlebot|Livelapbot|Mediapartners-Google|MJ12bot|MojeekBot|SemrushBot|SeznamBot|SummalyBot|VelenPublicWebCrawler|WebEMailExtrac|YandexBot|YisouSpider|