AI Search Crawler Inspector
Fetch any URL as ChatGPT, Claude, Perplexity, and Google. See exactly which AI bots can read your page — and which are blocked.
Results for
Why AI Search Crawler Inspector matters
If AI bots like ChatGPT, Claude, and Perplexity are blocked from reading your website, your content will never be cited in AI answers. Cloud firewalls, hidden JavaScript rendering, or overly strict robots.txt files can silently kill your AI visibility without ever showing up as broken in a normal web browser. Our tool fetches your URL across thirteen different AI agents simultaneously, revealing exactly who can access your content and who is getting shut out.
How it works
We analyze your robots.txt
We pull your live robots.txt file and parse it exactly like a real crawler would, testing each bot's user-agent token against your specific allow/disallow rules. Often, what webmasters think they configured differs from how bots actually interpret it.
We fetch your page as each bot
We send parallel live requests using the exact user-agent strings of 11 major AI crawlers, allowing you to see the exact response your server returns to them.
We compare the results
We aggregate the HTTP status codes, verify if real content was served, and cross-reference the robots.txt allowance—providing a clean, side-by-side view to instantly identify any blocked bots.
What it checks
HTTP status per bot
A 200 OK for a human browser might be a 403 Forbidden for an AI bot due to security layers. We expose these invisible failures side-by-side.
Real content presence
A 200 status isn't enough. We verify if the HTML actually contains readable text (like H1s and paragraphs), easily catching JavaScript-only blank pages or CDN bot challenges.
Robots.txt verdict
We evaluate your live robots.txt rules against your specific URL path to confirm if an AI agent is truly permitted to crawl it.
Training-only tokens
Tokens like Google-Extended govern data training usage rather than active web crawling. We report their exact directives directly from your robots.txt.