2026
Video-SVD: Efficient Video Diffusion via Orthogonal Basis Composition
ICML 2026poster
Video Diffusion Transformers (VDiTs) represent the state-of-the-art in video generation but are fundamentally constrained by the quadratic computational complexity of self-attention. To accelerate this critical computation, we analyze the pre-softmax matrix ($QK^T$) and reveal two key insights: (1) …