2026
Think-Then-Generate: Reasoning-Aware Text-to-Image Diffusion with LLM Encoders
ICML 2026poster
Recent progress in text-to-image (T2I) diffusion models (DMs) has enabled high-quality visual synthesis from diverse textual prompts. Yet, most existing T2I DMs, even those equipped with large language model (LLM)-based text encoders, remain text-pixel mappers -- they employ LLMs merely as text enco…