2026
MIDAS: Multi-Image Dispersion and Semantic Reconstruction for Jailbreaking MLLMs
ICLR 2026poster
Multimodal Large Language Models (MLLMs) have achieved remarkable performance but remain vulnerable to jailbreak attacks that can induce harmful content and undermine their secure deployment. Previous studies have shown that introducing additional inference steps, which disrupt security attention, c…