2025
DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification
CVPR 2025poster
In recent years, large transformer-based video encoder models have greatly advanced state-of-the-art performance on video classification tasks. However, these large models typically process videos by averaging embedding outputs from multiple clips over time to produce fixed-length representations. T…