2024
Free-ATM: Harnessing Free Attention Masks for Representation Learning on Diffusion-Generated Images
ECCV 2024poster
"This paper studies visual representation learning with diffusion-generated synthetic images. We start by uncovering that diffusion models’ cross-attention layers inherently provide annotation-free attention masks aligned with corresponding text inputs on generated images. We then investigate the pr…