2026
Mask-GCG: Are All Tokens in Adversarial Suffixes Necessary for Jailbreak Attacks?
ICASSP 2026poster
Jailbreak attacks on Large Language Models (LLMs) have demonstrated various successful methods whereby attackers manipulate models into generating harmful responses that they are designed to avoid. Among these, Greedy Coordinate Gradient (GCG) has emerged as a general and effective approach that opt…