How CitedScore Works
A live, repeatable process: crawl the site, generate real buyer questions, run them across four AI platforms, and score exactly what comes back. These are the constants the paid Report runs, published here because a methodology you can verify is worth more than one you have to trust.
A standard audit. Where a site's signals support only three buyer stages, CitedScore runs three personas and a proportionally smaller grid rather than inventing a fourth. The score weights never change.
How does CitedScore read your site?
CitedScore crawls your homepage first, then selects up to 14 more pages worth checking, one representative per site section rather than every URL on the domain. The homepage is evaluated once, from the same context every other check reads, so nothing gets fetched twice or scored on stale content.
Before choosing pages, the crawler requests robots.txt and skips every candidate page under a Disallow rule for User-agent: *. It reads Disallow lines only, so an Allow line does not reopen a blocked path. The homepage, the files it loads, and your sitemap are fetched either way, and if robots.txt errors or does not answer within five seconds, the sample goes ahead without it. Every sampled page runs the same checks:
- Title tag, present and 30 to 60 characters
- Meta description, present and 50 to 160 characters
- Exactly one H1 heading
- Structured data present, and which types
- Question-format headings
- Alt text on images
Canonical tags, HTTPS, sitemap presence and the other Technical SEO checks run once, on the homepage, and those homepage checks produce the Technical SEO Score. The sampled pages produce the Sitewide Coverage Sample and the Question Coverage section your Report shows. These are informational only and are never included in the Technical, AEO, or Overall scores.
How does CitedScore decide which questions to test?
Before a single platform is queried, CitedScore extracts a Company Profile from your site: pitch, category, offerings, price positioning, claimed differentiators, and the struggling moment your copy addresses. Nothing is invented. Every field is built only from signals actually present on the page.
From that profile it generates four buyer personas, one per stage of the buyer timeline, each carrying its own struggling moment, job to be done, and real search phrasing:
When a site's signals support only three distinct stages, CitedScore regenerates the full set once. If the second pass returns three again, the audit runs on three rather than inventing a fourth persona your site gave no evidence for. Three is the hard floor: below it the audit stops instead of guessing.
Each persona contributes three prompt runs. One question is asked twice, once in short conversational phrasing and once in long structured phrasing, as that persona's divergence pair. A second question is asked once. The pair exists because the two registers produce different AI answers, and a method that tested only one of them would miss half of what a buyer actually types.
A prompt is one question in one phrasing. An observation is what one platform returned for one prompt. The persona set is not the whole grid: eight further prompts run on every audit, four brand-neutral and four comparative, so the measurement is not built only from questions your own positioning suggested.
Which AI platforms does CitedScore test, and why those four?
ChatGPT, Google AI Overviews, Gemini, and Perplexity carry the overwhelming majority of buyer questions asked of AI today. Every audit checks all four: no platform add-ons, no per-platform pricing. They are not weighted equally, because your buyers are not evenly distributed across them.
| ChatGPT | 40% |
|---|---|
| Google AI Overviews | 35% |
| Gemini | 15% |
| Perplexity | 10% |
Per-platform scores are always shown alongside the blend, because a brand can be strong on one platform and invisible on another, and an average alone hides that. Twenty prompts run on every platform, producing 80 result observations per audit. The full grid appears in your Report, not a summary of it.
One retrieval detail: Google publishes no public API for AI Overviews, so those results are retrieved through DataForSEO, the search results provider disclosed on the sub-processors page. The other three platforms are queried through their own APIs.
How are the three scores calculated and weighted?
| Technical SEO | 30% |
|---|---|
| AEO Readiness | 30% |
| GEO Citation | 40% |
GEO receives the largest weight because direct AI visibility is the outcome the Report exists to measure. The GEO Citation Score blends the four platform scores using the usage-share weights above, and each platform score is calculated as:
Mentions carry 70% because a named recommendation drives revenue. Citations carry 30% because a source link has value even without the brand name attached. When a brand is mentioned, its position in the answer sets the weight applied to that mention:
| 1st position | 1.0 |
|---|---|
| 2nd position | 0.8 |
| 3rd position | 0.6 |
| 4th or later | 0.4 |
What is a ghost citation, and how does CitedScore detect one?
A mention names your brand in the visible answer text. A citation links your domain as a source. A ghost citation is a citation with no mention: the AI used your content, linked your URL, and never once said your name. Semrush and Kevin Indig's June 2026 research across 3,981 domain appearances found 61.7% of domain appearances were citations that never named the brand.
Every one of the 80 result observations is classified into exactly one of four states, using the same alias-matching logic on every platform:
- Both
- Named in the answer and cited as a source.
- Mentioned
- Named in the answer text, no source link.
- Cited only
- Source link, brand never named. A ghost citation.
- Absent
- Neither named nor cited.
The two failure modes need different fixes. A missing citation is an on-site problem: structure, schema, extractable content. A missing mention is a brand-recognition problem. CitedScore reports each lever separately, because treating them as one metric hides which one is actually broken.
What is the Maturity Ladder, and how is a rung earned?
A 0–100 score is precise, but it does not say what it means day to day. The CitedScore Maturity Ladder turns it into a five-rung read: your composite score proposes a rung, and your live AI results can pull it down from there.
Invisible
Not appearing in AI answers for your category
Overlooked
Named occasionally, rarely leading the answer
Mentioned
Consistently named across AI answers
Recommended
Named early and used as a source
Cited Authority
The default citation for your category
Your composite score sets the candidate rung. From there, evidence caps can only lower it, never raise it: a strong Technical score cannot buy a rung your live AI results do not support. The ladder is a read of your evidence; nothing flows from it back into the score. Every Report shows the specific, evidence-backed gate to the next rung.
What evidence sits behind every finding?
Every score traces back to a specific, named piece of evidence: the live prompt that produced it, the platform that returned it, and the actual response text. Nothing in the GEO section is a projection or an industry average applied to your domain.
Findings carry a tier. What CitedScore measured on your own audit is stated as fact, because a named field on your run supports it. Where a recommendation is standard AI-visibility practice rather than something measured on your site, the Report says so, stated with conviction as practice and never dressed up as a number this audit did not produce.
The audit is engine-assisted and engineer-verified: a diagnostic engine runs the full breadth of every check, and an engineer reviews every Report by hand. The engine gives coverage a manual review cannot match; the review gives judgment a machine cannot provide.
What changes between runs
AI answers vary from day to day. The same prompt, on the same platform, can name a different set of brands on a different day, so every audit is a dated reading of how these systems answer right now, and a baseline you can act on today.
This is why the audit is built on 80 observations rather than a handful. Every prompt runs across all four platforms, personas span the full buyer timeline, and the divergence pair asks the same question in both registers. The score is the pattern across that grid, not any single answer inside it. Results are aggregated across persona types and phrasing variants for the same reason: one query is an anecdote, eighty is a measurement.
Re-run an audit a week later and individual rows will move. A rung change or a shift of several points is a real signal worth acting on. Two or three points in either direction is the system breathing: normal movement around the same standing.
We ran this exact methodology on our own homepage and published the unedited result, including the fix it caught.
The parts people ask us to explain twice
If yours is not here, the eight sections above carry the full mechanism, and the research page shows the same method run end to end on our own site.
01What's the difference between being mentioned and being cited?
02How does the live visibility check work?
03Does improving my score actually change what AI says about me?
04Why is the GEO score weighted by platform?
05What does CitedScore actually show in the report?
06What is the CitedScore Maturity Ladder?
Run this exact methodology against your site.
One scan tells you where you're mentioned, where you're a ghost citation, and exactly what to fix.