Newsletter
Join the Community
Subscribe to our newsletter for the latest news and updates
Free tool to check if ChatGPT, Perplexity and Google AI can read and cite your site, backed by an original crawler-blocking census of the top 5,000 sites. No signup.
Submit your own product to reach creators and founders looking for the next tool to try.
AI Crawler Census checks whether a site's robots.txt blocks the crawlers that feed ChatGPT, Perplexity, and Google AI Overviews, and shows what a page looks like to those bots versus a normal browser -- paste a URL and it tells you which AI crawlers are allowed in, which are blocked, and why.
It's backed by an original census of the top 5,000 sites (Tranco list), collected 2026-09-07. Of the 4,999 sites checked, 2,771 have a robots.txt file at all. Among those:
The distinction matters because a training-crawler block and a search-crawler block have very different consequences: blocking GPTBot keeps your content out of OpenAI's model training, but blocking OAI-SearchBot or PerplexityBot means your pages simply cannot be cited in ChatGPT or Perplexity answers, even indirectly, no matter how good the content is.
The property also ships a free llms.txt generator/validator and a robots.txt linter, both usable without an account. No signup, no paid tier -- the whole thing runs client-side against the site you enter.