← Search

Chao Tong

3 accepted papers

2025

VidEvent: A Large Dataset for Understanding Dynamic Evolution of Events in Videos

AAAI 2025technical

Despite the significant impact of visual events on human cognition, understanding events in videos remains a challenging task for AI due to their complex structures, semantic hierarchies, and dynamic evolution. To address this, we propose the task of video event understanding that extracts event scr…

Cited by 0SourcePDFScholar
2024

H2GFormer: Horizontal-to-Global Voxel Transformer for 3D Semantic Scene Completion

AAAI 2024technical

3D Semantic Scene Completion (SSC) has emerged as a novel task in vision-based holistic 3D scene understanding. Its objective is to densely predict the occupancy and category of each voxel in a 3D scene based on input from either LiDAR or images. Currently, many transformer-based semantic scene comp…

2022

Learning Prototype via Placeholder for Zero-shot Recognition

IJCAI 2022poster

Zero-shot learning (ZSL) aims to recognize unseen classes by exploiting semantic descriptions shared between seen classes and unseen classes. Current methods show that it is effective to learn visual-semantic alignment by projecting semantic embeddings into the visual space as class prototypes. Ho…