← Search

Dawei Yan

5 accepted papers

2026

UVLM: Benchmarking Video Language Model for Underwater World Understanding

AAAI 2026technical

Recently, video-language models (VidLMs) have gained widespread attention and adoption. However, existing works primarily focus on terrestrial scenarios, overlooking the highly demanding application needs of underwater observation. To overcome this gap, we introduce UVLM, an under water observation

Cited by 0SourcePDFScholar
2025

MULiving: Towards Real-time Multi-User Survival State Monitoring Using Wearable RFID Tags

ICASSP 2025accepted

Human presence detection is crucial in various scenarios, from law enforcement surveillance to smart health-care systems. Traditional methods like cameras and acoustic signals face challenges such as privacy concerns, the need for line of sight (LoS), and susceptibility to environmental noise. This…

Cited by 0SourceScholar
2025

TG-LLaVA: Text Guided LLaVA via Learnable Latent Embeddings

AAAI 2025technical

Currently, inspired by the success of vision-language models (VLMs), an increasing number of researchers are focusing on improving VLMs and have achieved promising results. However, most existing methods concentrate on optimizing the connector and enhancing the language model component, while neglec…

Cited by 4SourcePDFScholar
2024

Low-Rank Rescaled Vision Transformer Fine-Tuning: A Residual Design Approach

CVPR 2024poster

Parameter-efficient fine-tuning for pre-trained Vision Transformers aims to adeptly tailor a model to downstream tasks by learning a minimal set of new adaptation parameters while preserving the frozen majority of pre-trained parameters. Striking a balance between retaining the generalizable represe…

2023

Efficient Adaptation of Large Vision Transformer via Adapter Re-Composing

NeurIPS 2023poster

The advent of high-capacity pre-trained models has revolutionized problem-solving in computer vision, shifting the focus from training task-specific models to adapting pre-trained models. Consequently, effectively adapting large pre-trained models to downstream tasks in an efficient manner has becom…