2023
Unified Discrete Diffusion for Simultaneous Vision-Language Generation
ICLR 2023poster
The recently developed discrete diffusion model performs extraordinarily well in generation tasks, especially in the text-to-image task, showing great potential for modeling multimodal signals. In this paper, we leverage these properties and present a unified multimodal generation model, which can p…