AIVIS
← All reference pages

AI bots: the complete user-agent list

Every AI answer engine uses a crawler with a declared user-agent. This list is generated directly from the same constant used by our scanning engine: it stays in sync automatically as new bots emerge.

User-agentEngine / purposeCategory
GPTBotOpenAI (ChatGPT) — trainingTraining
OAI-SearchBotOpenAI — ricerca/citazioni ChatGPTLive citation
ChatGPT-UserOpenAI — browsing live in ChatGPTLive citation
ClaudeBotAnthropic (Claude) — training/crawlingTraining
anthropic-aiAnthropic — crawling genericoTraining
Claude-UserAnthropic — browsing live in ClaudeLive citation
PerplexityBotPerplexity — crawling/citazioniLive citation
Perplexity-UserPerplexity — browsing liveLive citation
Google-ExtendedGoogle — training Gemini/AI OverviewsTraining
Applebot-ExtendedApple IntelligenceTraining
CCBotCommon Crawl (used to train many LLMs)Training
BytespiderByteDance (also used for AI training)Training
Meta-ExternalAgentMeta AITraining

"Training" gathers content to train future models; "Live citation" answers a specific query in real time, citing sources. Blocking one doesn't automatically block the other: they need to be managed separately in robots.txt.

Want to know which of these bots your site blocks today?

Check your robots.txt for free