← Search

Donguk Kim

2 accepted papers

2025

Don’t Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models

ACL 2025finding

Large Vision Language Models (LVLMs) demonstrate strong capabilities in visual understanding and description, yet often suffer from hallucinations, attributing incorrect or misleading features to images. We observe that LVLMs disproportionately focus on a small subset of image tokens—termed blind to…

2024

Flow-Assisted Motion Learning Network for Weakly-Supervised Group Activity Recognition

ECCV 2024poster

"Weakly-Supervised Group Activity Recognition (WSGAR) aims to understand the activity performed together by a group of individuals with the video-level label and without actor-level labels. We propose Flow-Assisted Motion Learning Network () for WSGAR, which consists of the motion-aware actor encode…

Cited by 1SourcePDFScholar