← Search

Kwanyoung Lee

3 accepted papers

2026

Adaptive Auxiliary Prompt Blending for Target-Faithful Diffusion Generation

CVPR 2026

Diffusion-based text-to-image (T2I) models have made remarkable progress in generating photorealistic and semantically rich images. However, when the target concepts lie in low-density regions of the training distribution, these models often produce semantically misaligned or structurally inconsiste

Cited by 0SourceScholar
2025

ScaleDiff: Higher-Resolution Image Synthesis via Efficient and Model-Agnostic Diffusion

NeurIPS 2025poster

Text-to-image diffusion models often exhibit degraded performance when generating images beyond their training resolution. Recent training-free methods can mitigate this limitation, but they often require substantial computation or are incompatible with recent Diffusion Transformer models. In this p…

Cited by 0SourceScholar
2025

VerbDiff: Text-Only Diffusion Models with Enhanced Interaction Awareness

CVPR 2025poster

Recent large-scale text-to-image diffusion models generate photorealistic images but often struggle to accurately depict interactions between humans and objects due to their limited ability to differentiate various interaction words.In this work, we propose VerbDiff to address the challenge of captu…