2025
One Head to Rule Them All: Amplifying LVLM Safety through a Single Critical Attention Head
NeurIPS 2025poster
Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities in tasks requiring multimodal understanding. However, recent studies indicate that LVLMs are more vulnerable than LLMs to unsafe inputs and prone to generating harmful content. Existing defense strategies primarily includ…