← Search

Hanqian Wu

3 accepted papers

2026

Synthetic Curriculum Reinforces Compositional Text-to-Image Generation

CVPR 2026

Text-to-Image (T2I) generation has long been an open problem, with compositional synthesis remaining particularly challenging. This task requires accurate rendering of complex scenes containing multiple objects that exhibit diverse attributes as well as intricate spatial and semantic relationships,

Cited by 0SourceScholar
2021

More than Text: Multi-modal Chinese Word Segmentation

ACL 2021short

Chinese word segmentation (CWS) is undoubtedly an important basic task in natural language processing. Previous works only focus on the textual modality, but there are often audio and video utterances (such as news broadcast and face-to-face dialogues), where textual, acoustic and visual modalities…

2021

Multi-modal Graph Fusion for Named Entity Recognition with Targeted Visual Guidance

AAAI 2021technical

Multi-modal named entity recognition (MNER) aims to discover named entities in free text and classify them into pre-defined types with images. However, dominant MNER models do not fully exploit fine-grained semantic correspondences between semantic units of different modalities, which have the poten…