2025
Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time
EMNLP 2025
Recently, Multimodal Large Language Models (MLLMs) have gained significant attention across various domains. However, their widespread adoption has also raised serious safety concerns.In this paper, we uncover a new safety risk of MLLMs: the output preference of MLLMs can be arbitrarily manipulated