← Search

Yifan Lan

1 accepted papers

2025

Phi: Preference Hijacking in Multi-modal Large Language Models at Inference Time

EMNLP 2025

Recently, Multimodal Large Language Models (MLLMs) have gained significant attention across various domains. However, their widespread adoption has also raised serious safety concerns.In this paper, we uncover a new safety risk of MLLMs: the output preference of MLLMs can be arbitrarily manipulated