2026
GeoShield: Safeguarding Geolocation Privacy from Vision-Language Models via Adversarial Perturbations
AAAI 2026technical
Vision-Language Models (VLMs) such as GPT-4o now demonstrate a remarkable ability to infer users
3 accepted papers
Vision-Language Models (VLMs) such as GPT-4o now demonstrate a remarkable ability to infer users
Multimodal large reasoning models (MLRMs) have advanced visual-textual integration, enabling sophisticated human-AI interaction. While prior work has exposed MLRMs to visual jailbreaks, it remains underexplored how their reasoning capabilities reshape the security landscape under adversarial inputs.
The field of few-shot learning (FSL) has shown promising results in scenarios where training data is limited, but its vulnerability to backdoor attacks remains largely unexplored. We first explore this topic by first evaluating the performance of the existing backdoor attack methods on few-shot lear…