← Search

Jonas Müller

3 accepted papers

2024

SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

ICLR 2024spotlight

We present Stable Diffusion XL (SDXL), a latent diffusion model for text-to-image synthesis. Compared to previous versions of Stable Diffusion, SDXL leverages a three times larger UNet backbone, achieved by significantly increasing the number of attention blocks and including a second text encoder.…

2024

Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

ICML 2024oral

Diffusion models create data from noise by inverting the forward paths of data towards noise and have emerged as a powerful generative modeling technique for high-dimensional, perceptual data such as images and videos. Rectified flow is a recent generative model formulation that connects data and no…

Cited by 1056SourcePDFScholar
2022

Retrieval-Augmented Diffusion Models

NeurIPS 2022accept

Novel architectures have recently improved generative image synthesis leading to excellent visual quality in various tasks. Much of this success is due to the scalability of these architectures and hence caused by a dramatic increase in model complexity and in the computational resources invested in…

Cited by 162SourcePDFScholar