← Search

Prateek Chhikara

1 accepted papers

2025

MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs

ICLR 2025poster

Multimodal Large Language Models (MLLMs) have experienced rapid progress in visual recognition tasks in recent years. Given their potential integration into many critical applications, it is important to understand the limitations of their visual perception. In this work, we study whether MLLMs can…