Methodology · Published In Full

How CitedScore Works

A live, repeatable process: crawl the site, generate real buyer questions, run them across four AI platforms, and score exactly what comes back. These are the constants the paid Report runs, published here because a methodology you can verify is worth more than one you have to trust.

4Buyer personas
20Prompts
4AI platforms
80Observations
30/30/40Score weights

A standard audit. Where a site's signals support only three buyer stages, CitedScore runs three personas and a proportionally smaller grid rather than inventing a fourth. The score weights never change.

01

How does CitedScore read your site?

CitedScore crawls your homepage first, then selects up to 14 more pages worth checking, one representative per site section rather than every URL on the domain. The homepage is evaluated once, from the same context every other check reads, so nothing gets fetched twice or scored on stale content.

Crawl budget
Pages sampled, homepage includedup to 15
Total crawl budget45s · 15 MB
Per-page cap1.5 MB
Per-page deadline30s

Before choosing pages, the crawler requests robots.txt and skips every candidate page under a Disallow rule for User-agent: *. It reads Disallow lines only, so an Allow line does not reopen a blocked path. The homepage, the files it loads, and your sitemap are fetched either way, and if robots.txt errors or does not answer within five seconds, the sample goes ahead without it. Every sampled page runs the same checks:

Canonical tags, HTTPS, sitemap presence and the other Technical SEO checks run once, on the homepage, and those homepage checks produce the Technical SEO Score. The sampled pages produce the Sitewide Coverage Sample and the Question Coverage section your Report shows. These are informational only and are never included in the Technical, AEO, or Overall scores.

02

How does CitedScore decide which questions to test?

Before a single platform is queried, CitedScore extracts a Company Profile from your site: pitch, category, offerings, price positioning, claimed differentiators, and the struggling moment your copy addresses. Nothing is invented. Every field is built only from signals actually present on the page.

From that profile it generates four buyer personas, one per stage of the buyer timeline, each carrying its own struggling moment, job to be done, and real search phrasing:

first thoughtMessy, symptom-based, conversational searches.
exploring optionsLearning casually, not yet committed to solving it.
evaluating optionsComparing in earnest, using specific category terms.
ready to decideNarrowed down, searching for trust signals.

When a site's signals support only three distinct stages, CitedScore regenerates the full set once. If the second pass returns three again, the audit runs on three rather than inventing a fourth persona your site gave no evidence for. Three is the hard floor: below it the audit stops instead of guessing.

Each persona contributes three prompt runs. One question is asked twice, once in short conversational phrasing and once in long structured phrasing, as that persona's divergence pair. A second question is asked once. The pair exists because the two registers produce different AI answers, and a method that tested only one of them would miss half of what a buyer actually types.

A prompt is one question in one phrasing. An observation is what one platform returned for one prompt. The persona set is not the whole grid: eight further prompts run on every audit, four brand-neutral and four comparative, so the measurement is not built only from questions your own positioning suggested.

How the 20 prompts decompose
Persona prompts (4 personas x 3 runs)12
Brand-neutral prompts4
Comparative prompts4
Prompts per audit20
Platforms each prompt runs on4
Observations per audit80
03

Which AI platforms does CitedScore test, and why those four?

ChatGPT, Google AI Overviews, Gemini, and Perplexity carry the overwhelming majority of buyer questions asked of AI today. Every audit checks all four: no platform add-ons, no per-platform pricing. They are not weighted equally, because your buyers are not evenly distributed across them.

Blend weight by usage share
ChatGPT40%
Google AI Overviews35%
Gemini15%
Perplexity10%

Per-platform scores are always shown alongside the blend, because a brand can be strong on one platform and invisible on another, and an average alone hides that. Twenty prompts run on every platform, producing 80 result observations per audit. The full grid appears in your Report, not a summary of it.

One retrieval detail: Google publishes no public API for AI Overviews, so those results are retrieved through DataForSEO, the search results provider disclosed on the sub-processors page. The other three platforms are queried through their own APIs.

04

How are the three scores calculated and weighted?

Overall = 30% Technical SEO + 30% AEO Readiness + 40% GEO Citation
Score blend
Technical SEO30%
AEO Readiness30%
GEO Citation40%

GEO receives the largest weight because direct AI visibility is the outcome the Report exists to measure. The GEO Citation Score blends the four platform scores using the usage-share weights above, and each platform score is calculated as:

platform = 100 × (0.7 × position-weighted mention rate + 0.3 × citation rate)

Mentions carry 70% because a named recommendation drives revenue. Citations carry 30% because a source link has value even without the brand name attached. When a brand is mentioned, its position in the answer sets the weight applied to that mention:

Position weighting
1st position1.0
2nd position0.8
3rd position0.6
4th or later0.4
05

What is a ghost citation, and how does CitedScore detect one?

A mention names your brand in the visible answer text. A citation links your domain as a source. A ghost citation is a citation with no mention: the AI used your content, linked your URL, and never once said your name. Semrush and Kevin Indig's June 2026 research across 3,981 domain appearances found 61.7% of domain appearances were citations that never named the brand.

Every one of the 80 result observations is classified into exactly one of four states, using the same alias-matching logic on every platform:

Mentioned + CitedBoth
Named in the answer and cited as a source.
MentionedMentioned
Named in the answer text, no source link.
Ghost citationCited only
Source link, brand never named. A ghost citation.
AbsentAbsent
Neither named nor cited.

The two failure modes need different fixes. A missing citation is an on-site problem: structure, schema, extractable content. A missing mention is a brand-recognition problem. CitedScore reports each lever separately, because treating them as one metric hides which one is actually broken.

06

How is share of voice measured?

Share of voice counts brands named, not words or sentiment. Your brand counts once per answer it appears in; each brand we identify as a competitor counts once per answer it is named in. The count runs only across answers the platform actually returned, so a query that errored or never triggered a response is excluded from the denominator rather than counted as a zero.

Queries that name your brand directly (“who is X”, “is X legit”, “X reviews”) are excluded from this metric on every audit as a standing rule: they almost always name you and almost never name a competitor, so counting them would flatter the figure. Below a minimum mention threshold, share of voice is not reported at all rather than published as a misleading small-sample number.

The exclusion moves share of voice and nothing else. Those queries still count in full toward the GEO score, because a navigational query you fail to win is still a real visibility miss. Share of voice is diagnostic, never a scoring input.

07

What is the Maturity Ladder, and how is a rung earned?

A 0–100 score is precise, but it does not say what it means day to day. The CitedScore Maturity Ladder turns it into a five-rung read: your composite score proposes a rung, and your live AI results can pull it down from there.

L1

Invisible

Not appearing in AI answers for your category

L2

Overlooked

Named occasionally, rarely leading the answer

L3

Mentioned

Consistently named across AI answers

L4

Recommended

Named early and used as a source

L5

Cited Authority

The default citation for your category

Your composite score sets the candidate rung. From there, evidence caps can only lower it, never raise it: a strong Technical score cannot buy a rung your live AI results do not support. The ladder is a read of your evidence; nothing flows from it back into the score. Every Report shows the specific, evidence-backed gate to the next rung.

08

What evidence sits behind every finding?

Every score traces back to a specific, named piece of evidence: the live prompt that produced it, the platform that returned it, and the actual response text. Nothing in the GEO section is a projection or an industry average applied to your domain.

Findings carry a tier. What CitedScore measured on your own audit is stated as fact, because a named field on your run supports it. Where a recommendation is standard AI-visibility practice rather than something measured on your site, the Report says so, stated with conviction as practice and never dressed up as a number this audit did not produce.

The audit is engine-assisted and engineer-verified: a diagnostic engine runs the full breadth of every check, and an engineer reviews every Report by hand. The engine gives coverage a manual review cannot match; the review gives judgment a machine cannot provide.

What changes between runs

AI answers vary from day to day. The same prompt, on the same platform, can name a different set of brands on a different day, so every audit is a dated reading of how these systems answer right now, and a baseline you can act on today.

This is why the audit is built on 80 observations rather than a handful. Every prompt runs across all four platforms, personas span the full buyer timeline, and the divergence pair asks the same question in both registers. The score is the pattern across that grid, not any single answer inside it. Results are aggregated across persona types and phrasing variants for the same reason: one query is an anecdote, eighty is a measurement.

Re-run an audit a week later and individual rows will move. A rung change or a shift of several points is a real signal worth acting on. Two or three points in either direction is the system breathing: normal movement around the same standing.

We ran this exact methodology on our own homepage and published the unedited result, including the fix it caught.

Questions About The Method

The parts people ask us to explain twice

If yours is not here, the eight sections above carry the full mechanism, and the research page shows the same method run end to end on our own site.

01What's the difference between being mentioned and being cited?
A mention means the AI names your brand in its answer: the buyer reads your name. A citation means the AI links your site as a source. They sound similar and they behave very differently. Semrush and Kevin Indig's June 2026 research across 3,981 domain appearances found that 61.7% of domain appearances were citations that never named the brand, where the AI used the site's content and never said who wrote it. CitedScore classifies every query result as Mentioned, Cited only, or Both, because the fix for a missing mention (brand recognition across the web) is different from the fix for a missing citation (extractable structure on your site).
02How does the live visibility check work?
CitedScore builds a buyer persona for each stage of the buyer timeline: four in the standard set, three when your site supports only three distinct stages, never fewer. From those personas it generates 20 targeted prompts. Twelve come from the personas themselves, three runs per persona: one question asked twice as that persona's divergence pair, once in short conversational phrasing and once in long structured phrasing, and a second question asked once. The remaining eight are brand-neutral and comparative queries. Each prompt runs across all four platforms simultaneously. We record whether your brand was named, whether your domain appeared as a source, your position among named results, and which competitor domains appeared instead. The full grid, 20 prompts across 4 platforms, appears in the report. A site that supports only three distinct buyer stages yields three personas and a proportionally smaller grid. Checking this manually across all four platforms for a single query takes 15 to 20 minutes. CitedScore runs all 80 result observations in the time it takes to get a coffee. Prompt results are aggregated across persona types and phrasing variants to reduce single-query variance. The pattern across 80 observations is more reliable than any individual result.
03Does improving my score actually change what AI says about me?
The two levers are real and separately measurable, and the report tells you which lever moves which result. On-site fixes (schema, structured content, llms.txt, question-based headings) move citations and extraction, and retrieval-driven platforms like ChatGPT and Perplexity are where that movement typically shows first. Mentions build through recognition: comparison content, community presence, entity verification, coverage that puts your name next to your category. CitedScore shows you exactly which signals are missing on each lever, in priority order. Fix them and re-run the audit; the per-platform scores show which lever moved.
04Why is the GEO score weighted by platform?
Because your buyers aren't evenly distributed. ChatGPT and Google AI Overviews reach far more people than the other platforms, so visibility there is worth more to your revenue. The blended GEO score weights each platform by usage share, which means the number tracks business impact instead of treating every platform as equal. The per-platform scores are always shown alongside it, because a brand can be strong on one platform and invisible on another, and an average alone would hide that.
05What does CitedScore actually show in the report?
The report shows the actual prompts we ran, the actual AI responses, which brands were named, which domains were cited, and your position or absence in each result. Every score in the report has evidence behind it: the live query that produced it, the platform that returned it, and the competitors that appeared instead of you. The GEO section is a full grid: 20 prompts across 4 platforms, each result classified as Mentioned, Cited only, Both, or Absent, grouped by buyer persona so you can see which types of buyers find you and which don't. The competitor gap table shows which domains are being named in your place and on which prompts. The recommendations section tells you which lever moves which result: on-site fixes that move citations, and off-site moves that build the brand recognition mentions require. Every recommendation carries a priority so you know what to fix first. Where the work can be sized, the plan shows the effort and who owns it.
06What is the CitedScore Maturity Ladder?
The CitedScore Maturity Ladder is a five-rung read of how visible and how trusted a brand is in AI-generated answers, from Invisible through Cited Authority. Your composite score sets the candidate rung, but the published rung comes from your own measured mention rate, citation rate, and mention position, not a self-assessment quiz. A strong score on paper cannot buy a rung your live AI results do not support. Every report shows the specific, evidence-backed gate to the next rung.
Ready To See Your Own Results

Run this exact methodology against your site.

One scan tells you where you're mentioned, where you're a ghost citation, and exactly what to fix.

CitedScore is in final build. Sign up now and get first access the day we launch. Reports are $79.

Protected by Cloudflare
How CitedScore Works: The Full AI Visibility Methodology | CitedScore