AAAI 2026technical0 citations

LAMP: Learning Universal Adversarial Perturbations for Multi-Image Tasks via Pre-trained Models

Alvi Md Ishmam, Najibul Haque Sarker, Zaber Ibn Abdul Hakim, Chris Thomas

Abstract

Multimodal Large Language Models (MLLMs) have achieved remarkable performance across vision-language tasks. Recent advancements allow these models to process multiple images as inputs. However, the vulnerabilities of multi-image MLLMs remain unexplored. Existing adversarial attacks focus on single-image settings and often assume a white-box threat model which is impractical in many real-world scenarios. This paper introduces LAMP, a black-box method for learning UAPs targeting multi-image MLLMs. LAMP applies an attention-based constraint that which prevents the model from effectively aggregating information across images. LAMP also introduces a novel cross-image contagious constraint that forces perturbed tokens to influence clean tokens to spread adversarial effects without requiring all inputs to be modified. Additionally, an index-attention suppression loss creates a robust position invariant attack. Experimental results show that LAMP outperforms SOTA baselines and achieves the highest attack success rates across multiple vision-language tasks.

BibTeX
@inproceedings{aaai2026_lamplearninguniv,
  title = {LAMP: Learning Universal Adversarial Perturbations for Multi-Image Tasks via Pre-trained Models},
  author = {Alvi Md Ishmam and Najibul Haque Sarker and Zaber Ibn Abdul Hakim and Chris Thomas},
  booktitle = {AAAI 2026},
  year = {2026}
}
LAMP: Learning Universal Adversarial Perturbations for Multi-Image Tasks via Pre-trained Models · AAAI 2026