← Search

Runqing Zhang

4 accepted papers

2026

Tackling Alignment Ambiguity in Person Retrieval through Conversational Attribute Mining

CVPR 2026

Text-to-Image Person Retrieval (TIPR) aims to retrieve pedestrian images with a given natural language description. It remains highly challenging due to the inherent ambiguity in cross-modal alignment: existing models often struggle to capture fine-grained correspondences, and their understanding of

Cited by 0SourcecodeScholar
2025

AMNS: Attention-Weighted Selective Mask and Noise Label Suppression for Text-to-Image Person Retrieval

ICASSP 2025accepted

Most existing text-to-image person retrieval methods usually assume that the training image-text pairs are perfectly aligned; however, the noisy correspondence(NC) issue (i.e., incorrect or unreliable alignment) exists due to poor image quality and labeling errors. Additionally, random masking augme…

Cited by 0SourceScholar
2025

FedDiT: Federated Learning by Distillation Token Enhanced Vision Transformer

ICASSP 2025accepted

Federated learning (FL) is a promising approach for privacy-preserving machine learning, enabling collaborative model training across distributed devices without sharing raw data. However, FL faces significant challenges due to the nonindependent and identically distributed (non-IID) nature of data…

Cited by 0SourceScholar
2025

Optimized Dynamic Watermarking for Audio DNNs with Adaptive Embedding and Boundary Sampling

ICASSP 2025accepted

The intensified concerns arising from the widespread adoption of deep learning have led to increased scrutiny of intellectual property protection in DNN models. Existing audio watermarking techniques, predominantly based on traditional signal processing methods, struggle to balance robustness, imper…

Cited by 0SourceScholar