← Search

Dongjun Hwang

2 accepted papers

2026

Enhancing Multi-Image Understanding through Delimiter Token Scaling

ICLR 2026poster

Large Vision-Language Models (LVLMs) achieve strong performance on single-image tasks, but their performance declines when multiple images are provided as input. One major reason is the cross-image information leakage, where the model struggles to distinguish information across different images. Exi…

Cited by 0SourcecodeScholar
2025

OVS Meets Continual Learning: Towards Sustainable Open-Vocabulary Segmentation

NeurIPS 2025poster

Open-Vocabulary Segmentation (OVS) aims to segment classes that are not present in the training dataset. However, most existing studies assume that the training data is fixed in advance, overlooking more practical scenarios where new datasets are continuously collected over time. To address this, we…

Cited by 0SourceScholar