AI Website Check
What does your site
tell AI crawlers?
One lookup against the record. We swept 449,070 domains on 6 September 2026 and wrote down what each one declared. If yours is in there, this is exactly what we read — not a score, not an estimate.
What this check is, and is not
We read robots.txt, the response headers and any published licence or Content-Signal directive, with a user agent that identifies itself as ours. We do not read your server logs, we cannot see which crawlers actually visited you, and a site charging privately through Cloudflare or a similar arrangement reads to us as undeclared — because from outside, it is.
Everything above comes from a single dated sweep. How a verdict is decided · The awkward questions
The five verdicts
Every domain in the record lands in exactly one of these. There is no sixth, and there is no “partly”.
Publishes a licence, or returned a price or HTTP 402.
robots.txt names an AI crawler and disallows it.
A Content-Signal preference, and nothing else.
Answered and declared nothing — so permitted by default.
Never answered us at all.
449,070 swept · 347,935 answered · 101,135 silent · sweep of 6 September 2026