Blocked_Bots_and_Crawler_Se.../apache-badbots.conf
oldkid 6afb7dc2b0 update ai.robots.txt
add robots.txt
add apache-badbots.conf
2025-04-26 06:57:32 +02:00

1 line
No EOL
1.1 KiB
Text

I2Bot|Ai2Bot\-Dolma|aiHitBot|Amazonbot|anthropic\-ai|Applebot|Applebot\-Extended|Brightbot\ 1\.0|Bytespider|CCBot|ChatGPT\-User|Claude\-Web|ClaudeBot|cohere\-ai|cohere\-training\-data\-crawler|Cotoyogi|Crawlspace|Diffbot|DuckAssistBot|FacebookBot|Factset_spyderbot|FirecrawlAgent|FriendlyCrawler|Google\-Extended|GoogleOther|GoogleOther\-Image|GoogleOther\-Video|GPTBot|iaskspider/2\.0|ICC\-Crawler|ImagesiftBot|img2dataset|imgproxy|ISSCyberRiskCrawler|Kangaroo\ Bot|meta\-externalagent|Meta\-ExternalAgent|meta\-externalfetcher|Meta\-ExternalFetcher|NovaAct|OAI\-SearchBot|omgili|omgilibot|Operator|PanguBot|Perplexity\-User|PerplexityBot|PetalBot|Scrapy|SemrushBot\-OCOB|SemrushBot\-SWA|Sidetrade\ indexer\ bot|TikTokSpider|Timpibot|VelenPublicWebCrawler|Webzio\-Extended|YouBot|AhrefsBot|Baiduspider|Barkrowler|Bingbot|BLEXBot|Bytedance|DotBot|EmailCollector|facebookcatalog|facebookexternalhit|fidget-spinner-bot|Franck the Fediverse Graph Crawler|Googlebot|Livelapbot|Mediapartners-Google|MJ12bot|SemrushBot|SeznamBot|VelenPublicWebCrawler|WebEMailExtrac|YandexBot|YisouSpider|