2026
Long-Text-to-Image Generation via Compositional Prompt Decomposition
ICLR 2026poster
While modern text-to-image models excel at generating images from intricate prompts, they struggle to capture the key details when the prompts are expanded into descriptive paragraphs. This limitation stems from the prevalence of short captions in their training data. Existing methods attempt to add…