This page is about one tool’s data: the AI Visibility Checker. For how these scores compare with the other two tools, the distributions and the percentile lookup, see SEO Score Benchmarks — the methodology and privacy notes are there too.
Which citability signals are missing
On the GEO (Generative Engine Optimization) and AEO (Answer Engine Optimization) side, what we measure is not ranking but citability: can a language model fetch the page, work out who wrote it, and extract an answer at passage level. The items below are the links in that chain; the percentage is the share of scans in which the check did not pass.
| Check | Did not pass | Weight |
|---|---|---|
| Author / byline | 43% | Partial |
| Organization / WebSite schema | 43% | Hard fail |
| llms.txt present | 40% | Partial |
| About page linked | 37% | Partial |
| Publish / update dates | 37% | Partial |
| Question-style headings | 30% | Partial |
| JSON-LD structured data | 20% | Hard fail |
| In-depth content (≥800 words) | 17% | Partial |
| Social / sameAs profiles linked | 13% | Partial |
| AI crawlers allowed | 10% | Hard fail |
| Brand mentioned consistently | 7% | Partial |
| Cites external sources | 7% | Partial |
The weakest AI-readiness area
Two problems run neck and neck: brand and entity strength and AI crawler access, each the weakest area in 37% of scans. The first is missing Organization and WebSite schema, absent sameAs profiles and an inconsistent brand name; the second is agents like GPTBot, ClaudeBot and PerplexityBot being blocked in robots.txt without anyone noticing. Both are cheap to fix and expensive to miss.
Plenty of schema — just not the schema that says who you are
One detail stands out: JSON-LD structured data is missing in only 20% of scans, but Organization/WebSite schema is missing in 43%. Sites do use schema — they just do not use the schema that describes themselves. To a language model, that is talking without ever saying who you are.
Frequently asked questions
What score do most sites get on an AI visibility check?
The median is 85/100 and 60% of scans reach 80+, but the distribution is bimodal. Look at the grade mix: 10 scans earned an A+ while 10 sat at a D. AI visibility is not gradual — either the basic machine-readable signals exist, or none of them do.
Which AI-readiness area is weakest?
Two run neck and neck: brand and entity strength, and AI crawler access, each the weakest area in 37% of scans. The first is missing Organization and WebSite schema, absent sameAs profiles and an inconsistent brand name; the second is agents like GPTBot, ClaudeBot and PerplexityBot being blocked in robots.txt without anyone noticing.
My site has JSON-LD — why is my AI visibility score still low?
Because using schema and using the schema that describes you are different things. JSON-LD is entirely absent in only 20% of scans, yet Organization/WebSite schema is missing in 43%. Sites add product, article or FAQ markup and skip the markup that states who they are. To a language model, that is talking without ever giving your name.
The sample, the methodology, the anonymization and the CC BY 4.0 licence are all set out in the methodology section on the hub.