← Search

Binglin Chen

1 accepted papers

2025

Cross-modal Causal Relation Alignment for Video Question Grounding

CVPR 2025highlight

Video question grounding (VideoQG) requires models to answer the questions and simultaneously infer the relevant video segments to support the answers. However, existing VideoQG methods usually suffer from spurious cross-modal correlations, leading to a failure to identify the dominant visual scenes…