
AI citation metrics confuse being cited with being chosen
Ahrefs' 75K-brand data says mentions get you named in AI answers; format data says comparison pages win the traffic. Citation-share KPIs measure the first and quietly claim the second.
The AI-visibility industry has picked its KPI, and I think it picked the wrong end of the funnel.
The data driving the playbook is real. Ahrefs ran correlation studies across 75,000 brands and found unlinked brand web mentions correlate with AI visibility at 0.664 Spearman, versus 0.218 for total backlinks — roughly a 3x gap, holding across ChatGPT, Google AI Mode, and AI Overviews. I verified those numbers against Ahrefs' own study pages. The consensus read wrote itself: get mentioned everywhere, get cited in AI answers, win.
Then the format data breaks the story in half. A Siege Media study of 116 B2B sites found "X vs Y" comparison pages the strongest predictor of AI search traffic at 0.65 Spearman — while "best X" listicles ranked last for traffic despite being among the most-cited formats. I have that figure only through a July 20, 2026 AISEO Weekly digest, not the primary study, so treat it as secondhand. But the shape of the finding is the interesting part: the formats models cite most and the formats that actually send you visitors are different formats.
Cited is not chosen
Hold both facts together and the standard KPI collapses. Mentions predict whether the model knows you. Format predicts whether the model's user picks you. A citation-count KPI measures the first and claims credit for the second — and the listicle-versus-comparison split shows those two things can move in opposite directions on the same site.
Paid slots complicate it further. Per the same digest — again secondhand — ads now appear in roughly 49% of US free ChatGPT replies as a labeled block. If that holds, the split becomes explicit: money buys placement below the answer, while citations still decide who gets named inside it. Bought placement and earned naming are now two separate markets, and a "citation share" dashboard tracks neither one's conversion.
One honest caveat before anyone rebuilds a strategy on this: every number here is correlational, measured on a moving target, and produced by SEO-tool vendors who sell measurement of the thing they measured. That is exactly why the confident "AEO playbooks" landing in inboxes this month deserve suspicion — they compress vendor correlations into causal advice inside one newsletter cycle.
The audit question that exposes it
At work, a client's marketing lead recently forwarded me an AEO audit proposal. The headline deliverable was citation share — "% of relevant ChatGPT answers that mention your brand." Nowhere in the scope was there a line connecting a citation to a session, or a session to pipeline. When I asked how they would attribute a closed deal to a mention, the answer was, honestly, that nobody attributes it — it is a visibility metric. That is the tell. Visibility metrics are fine as diagnostics and dangerous as KPIs, because budget follows the number on the dashboard.
My rule now: no AI-visibility spend gets approved on a single metric. It needs two, reported separately — naming rate (are we cited in answers for our category) and selection rate (do comparison-style pages that models surface actually convert visitors). The first tells you the model knows you exist. The second tells you the exposure earns money. Any vendor who reports one and implies the other is selling a vanity dashboard.
Steal this before your next marketing review: pull last quarter's traffic and assisted conversions for your listicle-style pages versus your comparison pages. If comparisons are already outperforming — which is what the Siege pattern predicts — shift the AEO budget toward building honest "us vs the alternative" pages, and demote citation counts to a diagnostic you glance at monthly. If a vendor pitches citation share as the KPI, make them add a selection metric to the contract or walk.
Being cited means the model knows you; being chosen means the user picked you — fund the metric that ends in pipeline, not the one that ends in a screenshot.


