About AI Crawler Checker
AI companies crawl the web to train models and to answer questions live. Use the AI crawler checker when you're deciding whether to block AI training, after a CDN or plugin has added AI rules for you, or when your site never shows up in AI search answers. Enter your site and click Check.
Your robots.txt is read for 25 AI crawlers, including GPTBot, OAI-SearchBot and ChatGPT-User from OpenAI, ClaudeBot and Claude-SearchBot from Anthropic, Google-Extended, Applebot-Extended, PerplexityBot, CCBot, Bytespider, meta-externalagent and Amazonbot. Each is labeled AI training, AI search or fetches for users, with allowed or blocked, the deciding rule and whether your file names it. The checker also looks for /llms.txt and /llms-full.txt and shows the start of llms.txt, and it spots noai or noimageai in meta robots or an X-Robots-Tag header. If AI search crawlers are blocked, you get a warning, because your pages could vanish from AI search answers. Write rules with the Robots.txt Generator and an llms.txt with the llms.txt Generator.
How to use AI Crawler Checker
- 1Enter your site
Type the domain or any page address on the site.
- 2Click Check
robots.txt, llms.txt, llms-full.txt and the home page are fetched.
- 3Read the table
Each AI crawler shows its purpose, allowed or blocked, and the deciding rule.
- 4Decide and adjust
Block training crawlers if you like, but keep AI search crawlers in if you want to be cited.
Why use Cubfile for this
- 25 AI crawlers
OpenAI, Anthropic, Google, Apple, Perplexity, Meta, Amazon, ByteDance and more.
- Purpose labels
AI training, AI search or fetches for users.
- llms.txt check
Finds llms.txt and llms-full.txt and shows how llms.txt begins.
- noai signals
Spots noai and noimageai in meta robots or X-Robots-Tag.