2026
Dual-Latent Memory Routing for Vision-Language Reasoning
ICML 2026spotlight
Multimodal large language models (MLLMs) have recently made strong progress in vision-language reasoning, yet their performance often degrades as generations grow longer. A key factor is that they frequently lose track of earlier visual evidence and intermediate constraints under a monolithic growin…