← Search

Yang Liu*

4 accepted papers

2024

Exploring Conditional Multi-Modal Prompts for Zero-shot HOI Detection

ECCV 2024poster

"Zero-shot Human-Object Interaction (HOI) detection has emerged as a frontier topic due to its capability to detect HOIs beyond a predefined set of categories. This task entails not only identifying the interactiveness of human-object pairs and localizing them but also recognizing both seen and unse…

2024

PiTe: Pixel-Temporal Alignment for Large Video-Language Model

ECCV 2024oral

"Fueled by the Large Language Models (LLMs) wave, Large Visual-Language Models (LVLMs) have emerged as a pivotal advancement, bridging the gap between image and text. However, video making it challenging for LVLMs to perform adequately due to the complexity of the relationship between language and s…

2024

Training-free Video Temporal Grounding using Large-scale Pre-trained Models

ECCV 2024poster

"Video temporal grounding aims to identify video segments within untrimmed videos that are most relevant to a given natural language query. Existing video temporal localization models rely on specific datasets for training, with high data collection costs, but exhibit poor generalization capability…