2024
CLAP: Isolating Content from Style through Contrastive Learning with Augmented Prompts
ECCV 2024poster
"Contrastive vision-language models, such as CLIP, have garnered considerable attention for various dowmsteam tasks, mainly due to the remarkable ability of the learned features for generalization. However, the features they learned often blend content and style information, which somewhat limits th…