← Search

Liu Yue

2 accepted papers

2025

Intend to Move: A Multimodal Dataset for Intention-Aware Human Motion Understanding

NeurIPS 2025poster

Human motion is inherently intentional, yet most motion modeling paradigms focus on low-level kinematics, overlooking the semantic and causal factors that drive behavior. Existing datasets further limit progress: they capture short, decontextualized actions in static scenes, providing little groundi…

Cited by 0SourceScholar
2024

ControlCap: Controllable Region-level Captioning

ECCV 2024poster

"Region-level captioning is challenged by the caption degeneration issue, which refers to that pre-trained multimodal models tend to predict the most frequent captions but miss the less frequent ones. In this study, we propose a controllable region-level captioning (ControlCap) approach, which intro…