← Search

Yukang Zhang

6 accepted papers

2026

FUSE: Frequency-domain Unification and Spectral Energy Alignment for Multi-modal Object Re-Identification

ICML 2026poster

Despite significant progress in multi-modal Re-Identification (ReID), existing methods tend to emphasize low-frequency cues. Consequently, they focus on attributes such as color, illumination, and coarse appearance, while overlooking mid- and high-frequency structures that encode geometric, textural…

Cited by 0SourceScholar
2026

Joint Implicit and Explicit Language Learning for Pedestrian Attribute Recognition

AAAI 2026technical

Pedestrian attribute recognition (PAR) has received increasing attention due to its wide application in video surveillance and pedestrian analysis. Some text-enhanced methods tackle this task by converting attributes into language descriptions to facilitate interactive learning between attributes an

Cited by 0SourcePDFScholar
2025

GSAlign: Geometric and Semantic Alignment Network for Aerial-Ground Person Re-Identification

NeurIPS 2025poster

Aerial-Ground person re-identification (AG-ReID) is an emerging yet challenging task that aims to match pedestrian images captured from drastically different viewpoints, typically from unmanned aerial vehicles (UAVs) and ground-based surveillance cameras. The task poses significant challenges due to…

Cited by 0SourceScholar
2025

MDReID: Modality-Decoupled Learning for Any-to-Any Multi-Modal Object Re-Identification

NeurIPS 2025spotlight

The challenge of inconsistent modalities in real-world applications presents significant obstacles to effective object re-identification (ReID). However, most existing approaches assume modality-matched conditions, significantly limiting their effectiveness in modality-mismatched scenarios. To overc…

Cited by 0SourceScholar
2024

RLE: A Unified Perspective of Data Augmentation for Cross-Spectral Re-Identification

NeurIPS 2024poster

This paper makes a step towards modeling the modality discrepancy in the cross-spectral re-identification task. Based on the Lambertain model, we observe that the non-linear modality discrepancy mainly comes from diverse linear transformations acting on the surface of different materials. From this…

2023

MRCN: A Novel Modality Restitution and Compensation Network for Visible-Infrared Person Re-identification

AAAI 2023technical

Visible-infrared person re-identification (VI-ReID), which aims to search identities across different spectra, is a challenging task due to large cross-modality discrepancy between visible and infrared images. The key to reduce the discrepancy is to filter out identity-irrelevant interference and ef…

Cited by 43SourcePDFScholar