← Search

Joe Penna

2 accepted papers

2024

SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

ICLR 2024spotlight

We present Stable Diffusion XL (SDXL), a latent diffusion model for text-to-image synthesis. Compared to previous versions of Stable Diffusion, SDXL leverages a three times larger UNet backbone, achieved by significantly increasing the number of attention blocks and including a second text encoder.…

2023

Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation

NeurIPS 2023poster

The ability to collect a large dataset of human preferences from text-to-image users is usually limited to companies, making such datasets inaccessible to the public. To address this issue, we create a web app that enables text-to-image users to generate images and specify their preferences. Using t…