What Crawlglass cannot do
This page exists because the rest of this category is full of promises nobody can keep. Read it before you pay us anything.
It cannot get you mentioned in ChatGPT
No tool can. Which sources an assistant names is decided inside a model we do not control and cannot see. Crawlglass makes sure you are reachable and readable — that is a precondition, not a guarantee.
It cannot tell you whether an AI has ever read your site
That would need your server logs. A scan sees your site from the outside, the same as a crawler does.
It cannot run JavaScript — and that is the point
We fetch your page with plain HTTP, because that is what most AI crawlers do. If your page looks fine in a browser and empty to us, that is a finding about your page, not a bug in ours.
It scans one page, not your whole site
The free scan looks at exactly the address you paste. Pro scans a list of pages on a schedule.
It reads robots.txt as the published rules say it should
Group selection by the most specific user-agent, longest match wins, Allow beats Disallow on a tie. Some crawlers are sloppier than that. We report what the rules mean, not what every bot does.
It stops at 2 MB, 5 redirects and 12 seconds
If your page exceeds those, we say so and report what we got. Crawlers have their own limits, usually tighter than ours.
Its score is an opinion, not a measurement
The individual findings are facts — your server really did return that status, your robots.txt really does say that. The number on top is our weighting of those facts. Treat the list as the product and the number as a summary.
It cannot see anything behind a login, a paywall or a bot wall
If a challenge page answers instead of your content, we report the wall. We will not try to get around it.
How we behave on your server
We identify honestly as CrawlglassBot/1.0 and never pretend to be another company’s crawler. A scan is a handful of requests: your page, your robots.txt and your llms.txt. We do not crawl onward from the page you give us.