2024
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
CVPR 2024poster
Large pre-trained Vision-Language Models (VLMs) like CLIP despite having remarkable generalization ability are highly vulnerable to adversarial examples. This work studies the adversarial robustness of VLMs from the novel perspective of the text prompt instead of the extensively studied model weight…