Activation Manipulation Attack: Penetrating and Harmful Jailbreak Attack Against Large Vision-Language Models
Recently, Large Vision-Language Models (LVLMs) have been demonstrated to be vulnerable to jailbreak attacks, highlighting the urgent need for further research to comprehensively identify and mitigate these threats. Unfortunately, existing jailbreak studies primarily focus on coarse-grained input man