← Search

Chongxin Li

1 accepted papers

2025

Attack as Defense: Safeguarding Large Vision-Language Models from Jailbreaking by Adversarial Attacks

EMNLP 2025

Adversarial vulnerabilities in vision-language models pose a critical challenge to the reliability of large language systems, where typographic manipulations and adversarial perturbations can effectively bypass language model defenses. We introduce Attack as Defense (AsD), the first approach to proa