2026
DAVSP: Safety Alignment for Large Vision-Language Models via Deep Aligned Visual Safety Prompt
AAAI 2026technical
Large Vision-Language Models (LVLMs) have achieved impressive progress across various applications but remain vulnerable to malicious queries. Existing safety alignment approaches typically fail to resist malicious queries while preserving utility on benign ones effectively. To address these challen