2025
SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization
ICCV 2025poster
This paper presents the Semantic-aWarE spatial-tEmporal Tokenizer (SweetTok), a novel video tokenizer to overcome the limitations in current video tokenization methods for compacted yet effective discretization. Unlike previous approaches that process flattened local visual patches via direct discre…