← Search

Zehao Dou

6 accepted papers

2025

Diffusion Transformer Captures Spatial-Temporal Dependencies: A Theory for Gaussian Process Data

ICLR 2025poster

Diffusion Transformer, the backbone of Sora for video generation, successfully scales the capacity of diffusion models, pioneering new avenues for high-fidelity sequential data generation. Unlike static data such as images, sequential data consists of consecutive data frames indexed by time, exhibit…

Cited by 3SourcePDFScholar
2025

Is Your Diffusion Model Actually Denoising?

NeurIPS 2025poster

We study the inductive biases of diffusion models with a conditioning-variable, which have seen widespread application as both text-conditioned generative image models and observation-conditioned continuous control policies. We observe that when these models are queried conditionally, their generati…

Cited by 0SourceScholar
2024

Theory of Consistency Diffusion Models: Distribution Estimation Meets Fast Sampling

ICML 2024poster

Diffusion models have revolutionized various application domains, including computer vision and audio generation. Despite the state-of-the-art performance, diffusion models are known for their slow sample generation due to the extensive number of steps involved. In response, consistency models have…

Cited by 4SourcePDFScholar