← Search

Trong-Thuan Nguyen

3 accepted papers

2025

HyperGLM: HyperGraph for Video Scene Graph Generation and Anticipation

CVPR 2025poster

Multimodal LLMs have advanced vision-language tasks but still struggle with understanding video scenes. To bridge this gap, Video Scene Graph Generation (VidSGG) has emerged to capture multi-object relationships across video frames. However, prior methods rely on pairwise connections, limiting their…

Cited by 1SourcePDFScholar
2024

CYCLO: Cyclic Graph Transformer Approach to Multi-Object Relationship Modeling in Aerial Videos

NeurIPS 2024poster

Video scene graph generation (VidSGG) has emerged as a transformative approach to capturing and interpreting the intricate relationships among objects and their temporal dynamics in video sequences. In this paper, we introduce the new AeroEye dataset that focuses on multi-object relationship modelin…

Cited by 3SourcePDFScholar
2024

HIG: Hierarchical Interlacement Graph Approach to Scene Graph Generation in Video Understanding

CVPR 2024poster

Visual interactivity understanding within visual scenes presents a significant challenge in computer vision. Existing methods focus on complex interactivities while leveraging a simple relationship model. These methods however struggle with a diversity of appearance situation position interaction an…

Cited by 14SourcePDFScholar