2026
Modal Aphasia: Can Unified Multimodal Models Describe Images From Memory?
ICLR 2026poster
We present *modal aphasia*, a systematic dissociation in which current unified multimodal models accurately memorize concepts visually but fail to articulate them in writing, despite being trained on images and text simultaneously. For one, we show that leading frontier models can generate near-perf…