2025
Diffuse Everything: Multimodal Diffusion Models on Arbitrary State Spaces
ICML 2025poster
Diffusion models have demonstrated remarkable performance in generating unimodal data across various tasks, including image, video, and text generation. On the contrary, the joint generation of multimodal data through diffusion models is still in the early stages of exploration. Existing approaches…