2025
Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models
ICLR 2025poster
Recent text-to-image diffusion models leverage cross-attention layers, which have been effectively utilized to enhance a range of visual generative tasks. However, our understanding of cross-attention layers remains somewhat limited. In this study, we introduce a mechanistic interpretability approac…