There is a failure mode in AI visibility reporting that almost every team has and almost nobody has named. You are in the answer. The engine used your page, linked it as a source, and built its response partly out of your work. And the reader finished reading without ever learning your company exists.
That is a ghost citation, and new data suggests it is the normal case rather than the exception. In a June 2026 study, Semrush and Kevin Indig ran 115 prompts across 14 countries and logged 3,981 domain appearances across four engines. Ghost citations, defined as a source link appearing while the brand name never appears in the answer text, accounted for 61.7% of them. Mentions with no citation took another 25.1%. Both together, the outcome every program is implicitly optimizing for, happened 13.2% of the time.
What ghost citations are
A citation is a machine-readable fact: the engine attached your URL to its answer. A mention is a human-readable fact: the reader saw your name. They come from different parts of the generation process, and they fail independently. Ghost citations are what you get when retrieval works and attribution does not.
| OUTCOME | SHARE OF APPEARANCES | WHAT THE READER EXPERIENCES |
|---|---|---|
| Ghost citation (cited, not named) | 61.7% | A footnote link they probably do not click |
| Mention only (named, not cited) | 25.1% | Your name in prose with no way to reach you |
| Both cited and named | 13.2% | The outcome everyone assumes they are buying |
Read the middle row again, because it is the stranger of the two failures. A quarter of the time the engine names a brand from its own parametric memory without linking anything. There is no click available. Nothing appears in your analytics. That appearance is invisible to every tracking method that starts from a referral, which is most of them, and it is one reason we keep arguing that the AI visibility reporting gap is a measurement problem before it is a content problem.
The engine split nobody accounts for
The aggregate hides the most actionable finding in the study. ChatGPT and Gemini behave close to opposite, and averaging them produces a number that describes neither.
How often a brand appearance includes a citation vs a name, by engine
ChatGPT cited 87% of the time and named the brand 20.7% of the time. Gemini inverted it: 21.4% cited, 83.7% named. Semrush's larger index of 126 million prompts points at the same structural difference from another angle, reporting ChatGPT citing an average of 15 sources per response against Gemini's 3. One engine is a bibliography. The other is a conversation that happens to remember brands.
That has a direct planning consequence. If your buyers research in ChatGPT, your ceiling is citations and your job is to be retrievable and correctly attributed in the source list. If they research in Gemini, citations are scarce and your job is to be the entity the model recalls unprompted, which is a brand and corroboration problem rather than a page-structure one. Running one strategy across both and reporting a blended visibility score is how programs end up unable to explain their own numbers.
“A blended AI visibility score across engines that disagree this violently is not a metric. It is an average of a bibliography and a conversation.”
Why ghost citations break your reporting
Three specific distortions follow from tracking one number when the underlying reality has two axes.
The third one is the expensive one. It is the mechanism by which a working AI program gets defunded: demand created in an untracked surface, converted through a tracked one, and attributed to the tracked one. Anyone who ran brand-versus-performance attribution arguments a decade ago has seen this movie.
Five plays that turn ghost citations into named mentions
Ghost citations are not a defect to eliminate. A citation is a real asset. The goal is to raise the share that also carry a name, and the levers are more specific than publish better content.
Play two has the strongest evidence behind it. The same study found comparative content produced 2.4x more brand mentions than informational content, which is consistent with what we have seen for a while: comparison pages earn the most citations and, it turns out, the most names as well. The buyer's first question is a comparison, and the answer to a comparison has to identify who is being compared.
The short-query finding in play three is the one most likely to change your roadmap. Short conversational prompts produced 30x to 50x more brand mentions than long ones. Long, heavily-specified prompts push the model toward synthesis, where names dissolve into a general answer. Short prompts push it toward recall, where names are the answer. Most enterprise prompt-tracking sets are full of long prompts because they read as more realistic, and they are systematically measuring the condition where brands are least likely to be named.
How to measure both numbers properly
The instrumentation is not complicated. It is just two columns where most teams keep one.
Report full_rate as the headline and the other three as diagnostics. A rising ghost rate with a flat full rate means your retrieval work is landing and your attribution work is not, which is a copy fix. A rising mention-only rate means parametric memory is improving, which is usually earned media finally compounding. This is the shape of the dashboards we build during reporting and analytics engagements, and the two-column split is the part clients keep after everything else changes.
One caveat on the numbers themselves: 115 prompts across 3,981 appearances is a real study but a modest one, and prompt selection drives results heavily in this kind of work. Treat 61.7% as directionally sound rather than precise, and run the same split on your own prompt set before you rebuild a roadmap around it. Semrush's own reporting notes that 45% of marketing leaders cannot accurately measure brand visibility in AI answers and only 9% have tools covering every relevant metric, so the honest position for most teams is that they do not yet know their own ghost rate.
Start here
Take twenty prompts your buyers actually run, weighted toward short ones. Run them across ChatGPT and Gemini. For each, mark two boxes: cited, and named. It takes an afternoon and produces the only version of this data that matters, which is yours.
Then look at the ghost rate specifically. If it is high, you have already won the hard half. Being retrieved is the part that depends on crawlability, structure, and trust, and the fix for the remaining half is largely editorial: write the claims so your name travels with them. If your ghost rate is low and your absent rate is high, the work is upstream, and it starts with the signals that earn a citation at all rather than with attribution phrasing. The full methodology behind the study is worth reading in Semrush's write-up of the AI visibility index, particularly if you operate in a category where the top three brands already hold most of the visibility. For B2B SaaS teams that concentration is the real competitive picture, and ghost citations are how you discover you are closer to it than your dashboard suggests.
See where you are cited today
A free snapshot audit of your rankings and AI citations before we ever talk.
Josh leads work at the intersection of SEO and generative engines at Something Inc., helping B2B brands get ranked and cited across every major AI engine.