2024
UniFashion: A Unified Vision-Language Model for Multimodal Fashion Retrieval and Generation
EMNLP 2024main
The fashion domain encompasses a variety of real-world multimodal tasks, including multimodal retrieval and multimodal generation. The rapid advancements in artificial intelligence generated content, particularly in technologies like large language models for text generation and diffusion models for…