2024
LLCP: Learning Latent Causal Processes for Reasoning-based Video Question Answer
ICLR 2024poster
Current approaches to Video Question Answering (VideoQA) primarily focus on cross-modality matching, which is limited by the requirement for extensive data annotations and the insufficient capacity for causal reasoning (e.g. attributing accidents). To address these challenges, we introduce a causal…