2024
Unleashing Text-to-Image Diffusion Prior for Zero-Shot Image Captioning
ECCV 2024poster
"Recently, zero-shot image captioning has gained increasing attention, where only text data is available for training. The remarkable progress in text-to-image diffusion model presents the potential to resolve this task by employing synthetic image-caption pairs generated by this pre-trained prior.…