← Search

Jinsung Lee

5 accepted papers

2026

Improving Target Presence and Plurality Recognition for Generalized Referring Image Segmentation

AAAI 2026technical

Generalized referring image segmentation (RIS) aims to segment regions in an image described by a natural language expression, handling not only single-target but also no- and multi-target scenarios. Previous approaches have proposed new components that enable a conventional RIS model to handle the

Cited by 0SourcePDFScholar
2026

Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model

CVPR 2026

World models provide a powerful framework for simulating environment dynamics conditioned on actions or instructions, enabling downstream tasks such as action planning or policy learning.Recent approaches leverage world models as learned simulators, but its application to decision-time planning rema

Cited by 0SourcecodeScholar
2024

Classification Matters: Improving Video Action Detection with Class-Specific Attention

ECCV 2024oral

"Video action detection (VAD) aims to detect actors and classify their actions in a video. We figure that VAD suffers more from classification rather than localization of actors. Hence, we analyze how prevailing methods form features for classification and find that they prioritize actor regions, ye…

Cited by 0SourcePDFScholar