The registry

24 agents tracked. 9 can move your score.

Every scan checks your robots.txt against this list, agent by agent. The tiers below are not a presentation choice — they are how the score is computed, and the other 15 agents are reported to you and deliberately not charged against your grade.

Tier 1 · 9 scored

Block one of these and you are out of AI answers

Answer-time fetchers and AI-search indexers with worldwide reach. Your robots_txt.allows_known_agents score is the share of these you allow, and nothing else feeds it.

AgentOperatorIntentrobots.txt token
ChatGPT-UserOpenAIAnswering a userChatGPT-User
OAI-SearchBotOpenAIAI search indexOAI-SearchBot
Claude-UserAnthropicAnswering a userClaude-User
Claude-SearchBotAnthropicAI search indexClaude-SearchBot
Perplexity-UserPerplexityAnswering a userPerplexity-User
PerplexityBotPerplexityAI search indexPerplexityBot
YouBotYou.comAI search indexYouBot
AmazonbotAmazonAnswering a userAmazonbot
ApplebotAppleAnswering a userApplebot
Tier 2 · 10 reported, not scored

Blocking a training crawler is a licensing decision

These collect text to train models. Whether you allow them is a business call about your content, not a readiness defect — blocking every one of them does not change how well an agent can use your site, so it does not move your grade. The scan tells you which are blocked and stops there.

AgentOperatorIntentrobots.txt token
GPTBotOpenAITraining dataGPTBot
ClaudeBotAnthropicTraining dataClaudeBotanthropic-ai
Google-ExtendedGoogleTraining dataGoogle-Extended
Applebot-ExtendedAppleTraining dataApplebot-Extended
BytespiderByteDanceTraining dataBytespider
CCBotCommon CrawlTraining dataCCBot
cohere-aiCohereTraining datacohere-ai
DiffbotDiffbotTraining dataDiffbot
Meta-ExternalAgentMetaTraining dataMeta-ExternalAgent
FacebookBotMetaTraining dataFacebookBot
Tier 3 · 5 checked conditionally

Regional engines, checked only if you say you want that market

Blocking Baidu is a defect for a shop selling into China and an entirely reasonable choice for a plumber in Ohio, and robots.txt alone cannot tell the two apart. So the scan reads what your own page declares — <html lang> and your hreflang alternates — and only reports these when your site says it wants those readers. Never scored: widening the denominator would penalise every site that never wanted those markets.

AgentOperatorIntentrobots.txt token
BaiduspiderBaidu · ChinaAI search indexBaiduspider
SogouSogou · ChinaAI search indexSogou web spiderSogou inst spiderSogou spider2
PetalBotHuawei · ChinaAI search indexPetalBot
YandexBotYandex · RussiaAI search indexYandexBot
YetiNaver · South KoreaAI search indexYeti

Which of these can reach your site?

Free, 30 seconds, no signup — verdict per agent, plus the fix for each block.

Every weight and threshold behind that verdict is published at /rubric, and our own measured error rate at /rubric/accuracy.

The 24 AI agents AgentSpeed tracks · AgentSpeed