Can AI assistants read your site?

An assistant can only name a business whose pages it can read. This checks both places a site gets blocked: the robots file, and the server itself.

who can read it example
2 of 5 cannot read this site
Blocked in the robots file, which almost nobody puts there on purpose.
ChatGPTblocked
Claudeblocked
Perplexityallowed
Geminiallowed
! An assistant cannot name what it cannot read.

This is one check out of many

The full scan adds security, speed, search, how the site reads on a phone, and roughly what year it was built.

Run the full scan → Read the site’s age too

Two checks, because a site can pass one and fail the other

  • The robots file. Read the way a crawler reads it, and we quote the exact line that decided.
  • The server. We ask for your homepage as each crawler. A refusal is a bot rule at the host or CDN, and it beats anything robots.txt says.

Should you block them?

It is a real choice. A publisher whose words are the product has good reason to. A plumber who wants the phone to ring almost never does, and usually has no idea the block is there.

Being readable is the floor, not a promise: it lets an assistant name you, it does not make it do so. We note an llms.txt if a site has one and score nothing for it, because no assistant requires it.