2023
CleanCLIP: Mitigating Data Poisoning Attacks in Multimodal Contrastive Learning
ICCV 2023oral
Multimodal contrastive pretraining has been used to train multimodal representation models, such as CLIP, on large amounts of paired image-text data. However, previous studies have revealed that such models are vulnerable to backdoor attacks. Specifically, when trained on backdoored examples, CLIP l…