← Search

Ngoc Dung Huynh

2 accepted papers

2026

VisRes Bench: On Evaluating the Visual Reasoning Capabilities of VLMs

CVPR 2026

Vision-Language Models (VLMs) have achieved remarkable progress across tasks such as visual question answering and image captioning. Yet, the extent to which these models perform visual reasoning as opposed to relying on linguistic priors remains unclear. To address this, we introduce VisRes Bench,

Cited by 0SourcecodeScholar
2025

Vision-Language Models Can't See the Obvious

ICCV 2025poster

We present Saliency Benchmark (SalBench), a novel benchmark designed to assess the capability of Large Vision-Language Models (LVLM) in detecting visually salient features that are readily apparent to humans, such as a large circle amidst a grid of smaller ones. This benchmark focuses on low-level f…

Cited by 0SourcePDFScholar