2026
Action-Guided Attention for Video Action Anticipation
ICLR 2026poster
Anticipating future actions in videos is challenging, as the observed frames provide only evidence of past activities, requiring the inference of latent intentions to predict upcoming actions. Existing transformer-based approaches, which rely on dot-product attention over pixel representations, ofte…