The AI crawler directory
61 known AI crawlers that may visit your site — who runs them, what they do with your content, and whether a robots.txt Disallow actually stops them. Every profile links back to blocking instructions and the free audit tool.
AI search crawlers (10)
AI assistant fetchers (11)
AI agent browsers (9)
AI training data crawlers (31)
GPTBotOpenAI
ClaudeBotAnthropic
anthropic-aiAnthropic
Google-ExtendedGoogle
GoogleOtherGoogle
GoogleOther-ImageGoogle
GoogleOther-VideoGoogle
Google-CloudVertexBotGoogle
Applebot-ExtendedApple
BytespiderByteDance · ⚠ ignores robots.txt
TikTokSpiderByteDance
CCBotCommon Crawl
FacebookBotMeta
meta-externalagentMeta
BedrockbotAmazon
DeepSeekBotDeepSeek · ⚠ ignores robots.txt
MistralAI-TrainingMistral AI
cohere-training-data-crawlerCohere
DiffbotDiffbot
AI2BotAllen Institute for AI
Ai2Bot-DolmaAllen Institute for AI
img2datasetimg2dataset project
LAIONDownloaderLAION · ⚠ ignores robots.txt
ImagesiftBotHive AI
ICC-CrawlerNICT (Japan)
ISSCyberRiskCrawlerISS-Corporate · ⚠ ignores robots.txt
FriendlyCrawlerUnknown
PanguBotHuawei
YandexAdditionalYandex
OmgilibotWebz.io
OmgiliWebz.io
Check your site against all 61 crawlers
Paste your robots.txt and get a per-crawler verdict, a score, and a fix list — free, no signup, runs in your browser.