2026
Mechanistic Dissection of Cross-Attention Subspaces in Text-to-Image Diffusion Models
AAAI 2026technical
Text-to-image diffusion models utilize cross-attention to integrate textual information into the visual latent space, yet the transformation from text embeddings to latent features remains largely unexplored. We provide a mechanistic analysis of the output-value (OV) circuits within cross-attention