← Search

Huanzhang Dou

7 accepted papers

2025

IDEA-Bench: How Far are Generative Models from Professional Designing?

CVPR 2025poster

Recent advancements in image generation models enable the creation of high-quality images and targeted modifications based on textual instructions. Some models even support multimodal complex guidance and demonstrate robust task generalization capabilities. However, they still fall short of meeting…

2024

ScanFormer: Referring Expression Comprehension by Iteratively Scanning

CVPR 2024poster

Referring Expression Comprehension (REC) aims to localize the target objects specified by free-form natural language descriptions in images. While state-of-the-art methods achieve impressive performance they perform a dense perception of images which incorporates redundant visual regions unrelated t…

Cited by 10SourcePDFScholar
2023

GaitGCI: Generative Counterfactual Intervention for Gait Recognition

CVPR 2023poster

Gait is one of the most promising biometrics that aims to identify pedestrians from their walking patterns. However, prevailing methods are susceptible to confounders, resulting in the networks hardly focusing on the regions that reflect effective walking patterns. To address this fundamental proble…

Cited by 59SourcePDFScholar
2023

Language Adaptive Weight Generation for Multi-Task Visual Grounding

CVPR 2023poster

Although the impressive performance in visual grounding, the prevailing approaches usually exploit the visual backbone in a passive way, i.e., the visual backbone extracts features with fixed weights without expression-related hints. The passive perception may lead to mismatches (e.g., redundant and…

2023

Referring Expression Comprehension Using Language Adaptive Inference

AAAI 2023technical

Different from universal object detection, referring expression comprehension (REC) aims to locate specific objects referred to by natural language expressions. The expression provides high-level concepts of relevant visual and contextual patterns, which vary significantly with different expressions…

Cited by 20SourcePDFScholar
2022

Adaptive Cross-Domain Learning for Generalizable Person Re-identification

ECCV 2022poster

"Domain Generalizable Person Re-Identification (DG-ReID) is a more practical ReID task that is trained from multiple source domains and tested on the unseen target domains. Most existing methods are challenged for dealing with the shared and specific characteristics among different domains, which is…

2022

MetaGait: Learning to Learn an Omni Sample Adaptive Representation for Gait Recognition

ECCV 2022poster

"Gait recognition, which aims at identifying individuals by their walking patterns, has recently drawn increasing research attention. However, gait recognition still suffers from the conflicts between the limited binary visual clues of the silhouette and numerous covariates with diverse scales, whic…

Cited by 46SourcePDFScholar