← Search

Jungmin Ko

2 accepted papers

2026

DOS: Directional Object Separation in Text Embeddings for Multi-Object Image Generation

AAAI 2026technical

Recent progress in text-to-image (T2I) generative models has led to significant improvements in generating high-quality images aligned with text prompts. However, these models still struggle with prompts involving multiple objects, often resulting in object neglect or object mixing. Through extensiv

Cited by 0SourcePDFScholar
2025

Cross-Attention Head Position Patterns Can Align with Human Visual Concepts in Text-to-Image Generative Models

ICLR 2025poster

Recent text-to-image diffusion models leverage cross-attention layers, which have been effectively utilized to enhance a range of visual generative tasks. However, our understanding of cross-attention layers remains somewhat limited. In this study, we introduce a mechanistic interpretability approac…