← Search

Shengming Yuan

5 accepted papers

2026

Understanding and Mitigating Token-Pruning-Induced Vulnerabilities in VLMs

ICML 2026poster

Token-Pruning accelerates Vision-Language Models by removing redundant visual tokens, yet its safety implications remain underexplored. In this work, we present the first comprehensive safety evaluation of Token-Pruning mechanism and find that: Most pruning strategies significantly degrade safety as…

Cited by 0SourceScholar
2025

FlexAC: Towards Flexible Control of Associative Reasoning in Multimodal Large Language Models

NeurIPS 2025poster

Multimodal large language models (MLLMs) face an inherent trade-off between faithfulness and creativity, as different tasks require varying degrees of associative reasoning. However, existing methods lack the flexibility to modulate this reasoning strength, limiting MLLMs' adaptability across factua…

Cited by 0SourcecodeScholar
2025

SafePTR: Token-Level Jailbreak Defense in Multimodal LLMs via Prune-then-Restore Mechanism

NeurIPS 2025poster

By incorporating visual inputs, Multimodal Large Language Models (MLLMs) extend LLMs to support visual reasoning. However, this integration also introduces new vulnerabilities, making MLLMs susceptible to multimodal jailbreak attacks and hindering their safe deployment. Existing defense methods, inc…

Cited by 0SourceScholar
2024

Any Target Can be Offense: Adversarial Example Generation via Generalized Latent Infection

ECCV 2024poster

"Targeted adversarial attack, which aims to mislead a model to recognize any image as a target object by imperceptible perturbations, has become a mainstream tool for vulnerability assessment of deep neural networks (DNNs). Since existing targeted attackers only learn to attack known target classes,…

2022

Natural Color Fool: Towards Boosting Black-box Unrestricted Attacks

NeurIPS 2022accept

Unrestricted color attacks, which manipulate semantically meaningful color of an image, have shown their stealthiness and success in fooling both human eyes and deep neural networks. However, current works usually sacrifice the flexibility of the uncontrolled setting to ensure the naturalness of adv…