left arrow icon

No account neededWe read your robots.txt once and keep nothing

AI crawler checker

See whether ChatGPT, Claude, Perplexity and Google's AI features may read a page under your robots.txt, and which crawlers only collect training data.

Free, no account. We read your robots.txt once and keep nothing.

Your reading appears here

Enter a page of your site to see what its robots.txt tells each AI crawler. Or see a finished reading first.

Three kinds of AI crawler

AI search
Index the pages AI answers cite: OAI-SearchBot for ChatGPT search, Claude-SearchBot, PerplexityBot, and Googlebot, whose index feeds Google’s AI Overviews.
Live reads
Fetch one page when a person asks about it: ChatGPT-User, Claude-User, Perplexity-User. OpenAI and Perplexity say robots.txt may not apply to theirs.
Training
Collect pages to train models: GPTBot, ClaudeBot, Google-Extended, Applebot-Extended, meta-externalagent and Common Crawl’s CCBot.

Stay in AI answers, stay out of training

User-agent: GPTBot
User-agent: ClaudeBot
User-agent: Google-Extended
User-agent: Applebot-Extended
User-agent: meta-externalagent
User-agent: CCBot
Disallow: /

User-agent: *
Allow: /
  1. Crawlers named together share one group. All six read the Disallow under them.
  2. Every crawler not named reads the * group. The search crawlers are not named, so they may fetch everything.
  3. A crawler with its own group ignores the * group. Give a named crawler every rule it needs, the private paths included.

Questions

Does blocking GPTBot take my site out of ChatGPT?

No. OpenAI documents GPTBot, which collects training data, and OAI-SearchBot, which powers ChatGPT search, as separate crawlers with separate robots.txt rules.

Does Cloudflare block AI crawlers on my site?

It can, in two ways. It can add rules to your robots.txt, which this check reads and labels, and it can refuse crawlers at its firewall, which no outside check can see. Your Cloudflare dashboard shows both.

Do I need an llms.txt file?

Not for crawler access: llms.txt is not a robots.txt rule, and it allows or blocks nothing. Google says its Search does not use it.

Do you keep the sites I check?

No. Our server fetches the robots.txt once, sends you the reading, and stores neither the address nor the result.

Awiser is the builder network where founders show their skills, find collaborators, and grow their projects with an AI copilot. More mini-apps