2026
How Do Medical MLLMs Fail? A Study on Visual Grounding in Medical Images
ICLR 2026poster
Generalist multimodal large language models (MLLMs) have achieved impressive performance across a wide range of vision-language tasks. However, their performance on medical tasks—particularly in zero-shot settings where generalization is critical—remains suboptimal. A key research gap is the limited…