← Search

Zhongqi Wang

3 accepted papers

2026

Complementary Subspace Low-Rank Adaptation of Vision-Language Models for Few-Shot Classification

ICASSP 2026oral

Vision language model (VLM) has been designed for large scale image-text alignment as a pretrained foundation model. For downstream few shot classification tasks, parameter efficient fine-tuning (PEFT) VLM has gained much popularity in the computer vision community. PEFT methods like prompt tuning a…

Cited by 0SourcePDFScholar
2025

Dysca: A Dynamic and Scalable Benchmark for Evaluating Perception Ability of LVLMs

ICLR 2025poster

Currently many benchmarks have been proposed to evaluate the perception ability of the Large Vision-Language Models (LVLMs). However, most benchmarks conduct questions by selecting images from existing datasets, resulting in the potential data leakage. Besides, these benchmarks merely focus on evalu…

2024

T2IShield: Defending Against Backdoors on Text-to-Image Diffusion Models

ECCV 2024poster

"While text-to-image diffusion models demonstrate impressive generation capabilities, they also exhibit vulnerability to backdoor attacks, which involve the manipulation of model outputs through malicious triggers. In this paper, for the first time, we propose a comprehensive defense method named T2…