Crawler directory

The crawlers that decide whether you show up in search and AI answers — what they do, how they identify themselves and what blocking them costs you.

CrawlerOperatorPurposerobots.txt token
GooglebotGoogleSearch indexing (also classified as AI training by Cloudflare)Googlebot
GPTBotOpenAICollecting data to train OpenAI modelsGPTBot
OAI-SearchBotOpenAIIndexing for ChatGPT searchOAI-SearchBot
ChatGPT-UserOpenAIFetching pages when a ChatGPT user asks for itChatGPT-User
ClaudeBotAnthropicCollecting data to train Anthropic modelsClaudeBot
Claude-SearchBotAnthropicIndexing to improve Claude search resultsClaude-SearchBot
Claude-UserAnthropicFetching pages when a Claude user asks for itClaude-User
PerplexityBotPerplexityIndexing for Perplexity search answersPerplexityBot
Perplexity-UserPerplexityFetching pages when a Perplexity user asks for itPerplexity-User
BingbotMicrosoftBing search indexing (also feeds Copilot and many AI search products)bingbot
ApplebotAppleSearch for Siri, Spotlight and Safari suggestionsApplebot
Google-ExtendedGooglerobots.txt token controlling use for Gemini training and groundingGoogle-Extended
Crawler directory: user-agents, purpose and robots.txt tokens · BotCanary