← Search

Fangfang Wang

4 accepted papers

2025

Few-Shot Incremental Multi-modal Learning via Touch Guidance and Imaginary Vision Synthesis

IJCAI 2025

Multimodal perception, which integrates vision and touch, is increasingly demonstrating its significance in domains such as embodied intelligence and human-computer interaction. However, in open-world scenarios, multimodal data streams face significant challenges, including catastrophic forgetting a

2024

Self-Distilled Dynamic Fusion Network for Language-Based Fashion Retrieval

ICASSP 2024accepted

In the domain of language-based fashion image retrieval, pinpointing the desired fashion item using both a reference image and its accompanying textual description is an intriguing challenge. Existing approaches lean heavily on static fusion techniques, intertwining image and text. Despite their com…

Cited by 0SourceScholar
2020

BANet: Bidirectional Aggregation Network With Occlusion Handling for Panoptic Segmentation

CVPR 2020oral

Panoptic segmentation aims to perform instance segmentation for foreground instances and semantic segmentation for background stuff simultaneously. The typical top-down pipeline concentrates on two key issues: 1) how to effectively model the intrinsic interaction between semantic segmentation and in…

Cited by 92PDFcodeScholar
2018

Geometry-Aware Scene Text Detection With Instance Transformation Network

CVPR 2018poster

Localizing text in the wild is challenging in the situations of complicated geometric layout of the targets like random orientation and large aspect ratio. In this paper, we propose a geometry-aware modeling approach tailored for scene text representation with an end-to-end learning scheme. In our a…

Cited by 111SourcePDFScholar