Skip to main content

CometWeb tools

AI visibility checker

A free check of one page for AI crawlers · robots.txt for 14 crawlers, text without JavaScript, directives and schema

FreeNo account required

Our server fetches the URL, plus that domain's robots.txt and llms.txt, so it reaches our logs. Private and local addresses are rejected, including after a redirect.

This tool checks whether AI crawlers can read and quote the page. It does not measure citations or rankings in ChatGPT, Perplexity, Gemini or AI Overviews, because that data is not public.

This AI visibility checker reads one public page the way an AI crawler gets it: the server HTML, robots.txt and the indexing directives, skipping JavaScript. It checks whether AI search crawlers such as OAI-SearchBot, Claude-SearchBot and PerplexityBot may fetch the page and how much text is left without JavaScript. It also checks whether noindex or nosnippet limits quoting, and which structured data and llms.txt the site has. The result shows readiness to be read. Citations in ChatGPT, Perplexity or Google stay out of reach for any outside tool, because those systems keep that data to themselves.

What to fix first

Start with robots.txt. A single Disallow: / under the wrong user agent hides the page from an AI search engine. Allow the search crawlers you care about, and decide about training crawlers separately. You can test each rule in our robots.txt tester.

Then check the text without JavaScript. Open the page source (Ctrl+U) and search for a sentence from your main content. If it is missing there, a crawler that skips scripts misses it too. Server-side rendering or prerendering fixes that for the whole site.

Directives and schema come next. A noindex or nosnippet left over from a staging setup is a common reason a page is missing from results. Structured data and an llms.txt file help machines read the page, with no promise of citations.

This is one page at one moment. To check every page and repeat the check after each release, see AI search readiness in CometWeb Insight.

Check several kinds of page

robots.txt rules and rendering often differ between templates. Check the home page, one article or product and a category page. If only one template has little text in the server HTML, the fix belongs to that template, not to the whole site.

Google Search Central · AI features and your website ↗

Terms in the result

robots.txtRobots Exclusion Protocol (RFC 9309)
A file with rules on which pages each crawler may fetch.
User-agentCrawler name
The token a crawler looks for to find its rules in robots.txt.
SSRServer-side rendering
The server sends finished HTML with the content, without waiting for JavaScript.
nosnippetRobots directive
Turns off the snippet in results, and in Google also use in AI Overviews.
JSON-LDJSON for Linked Data
schema.org structured data written in a script tag.
This tool patches one problem at a time

One page is a sample. The SEO module in Insight checks AI crawler access in robots.txt and the other signals across the whole site, then repeats the check after a release.