← Search

Po-Yuan Mao

1 accepted papers

2025

VSC: Visual Search Compositional Text-to-Image Diffusion Model

ICCV 2025poster

Text-to-image diffusion models have shown impressive capabilities in generating realistic visuals from natural-language prompts, yet they often struggle with accurately binding attributes to corresponding objects, especially in prompts containing multiple attribute-object pairs. This challenge prima…

Cited by 0SourcePDFScholar