← Search

Songsong Duan

3 accepted papers

2025

DIH-CLIP: Unleashing the Diversity of Multi-Head Self-Attention for Training-Free Open-Vocabulary Semantic Segmentation

ICCV 2025poster

Recent Training-Free Open-Vocabulary Semantic Segmentation (TF-OVSS) leverages a pre-training vision-language model to segment images from open-set visual concepts without training and fine-tuning. The key of TF-OVSS is to improve the local spatial representation of CLIP by leveraging self-correlati…

2025

Dual Information Purification for Lightweight SAR Object Detection

AAAI 2025technical

Synthetic aperture radar (SAR) object detection requires accurate identification and localization of targets at various scales within SAR images. However, background clutter and speckle noise can obscure key features and mislead the knowledge distillation process. To address these challenges, we int…

Cited by 1SourcePDFScholar
2025

Multi-Label Prototype Visual Spatial Search for Weakly Supervised Semantic Segmentation

CVPR 2025highlight

Existing Weakly Supervised Semantic Segmentation (WSSS) relies on the CNN-based Class Activation Map (CAM) and Transformer-based self-attention map to generate class-specific masks for semantic segmentation. However, CAM and self-attention maps usually cause incomplete segmentation due to classifica…

Cited by 0SourcePDFScholar