2025
Emergent Temporal Correspondences from Video Diffusion Transformers
NeurIPS 2025poster
Recent advancements in video diffusion models based on Diffusion Transformers (DiTs) have achieved remarkable success in generating temporally coherent videos. Yet, a fundamental question persists: how do these models internally establish and represent temporal correspondences across frames? We int…