Omni-Attack: Adversarial Attacks on Open-Ended VQA in Black-Box Multimodal LLMs
Multimodal large language models (MLLMs) have achieved remarkable success across diverse applications, from autonomous driving to document understanding. As these models are deployed in safety-critical contexts, understanding their adversarial robustness becomes crucial. However, current evaluations