← Search

Youngsun Lim

2 accepted papers

2026

What "Not" to Detect: Negation-Aware VLMs via Structured Reasoning and Token Merging

ICLR 2026poster

State-of-the-art vision-language models (VLMs) suffer from a critical failure in understanding negation, often referred to as affirmative bias. This limitation is particularly severe in described object detection (DOD) tasks. To address this, we propose two primary contributions: (1) a new dataset p…

Cited by 0SourceScholar
2025

Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering

AAAI 2025technical

Despite the impressive success of text-to-image (TTI) models, existing studies overlook the issue of whether these models accurately convey factual information. In this paper, we focus on the problem of image hallucination, where images created by TTI models fail to faithfully depict factual content…