← Search

Kok Pin Ng

1 accepted papers

2026

How Do Medical MLLMs Fail? A Study on Visual Grounding in Medical Images

ICLR 2026poster

Generalist multimodal large language models (MLLMs) have achieved impressive performance across a wide range of vision-language tasks. However, their performance on medical tasks—particularly in zero-shot settings where generalization is critical—remains suboptimal. A key research gap is the limited…

Cited by 0SourceScholar