← Search

Fangsheng Weng

4 accepted papers

2026

GarmentGPT: Compositional Garment Pattern Generation via Discrete Latent Tokenization

ICLR 2026poster

Apparel is a fundamental component of human appearance, making garment digitalization critical for digital human creation. However, sewing pattern creation traditionally relies on the intuition and extensive experience of skilled artisans. This manual bottleneck significantly hinders the scalability…

Cited by 0SourcecodeScholar
2024

Tailored Visions: Enhancing Text-to-Image Generation with Personalized Prompt Rewriting

CVPR 2024poster

Despite significant progress in the field it is still challenging to create personalized visual representations that align closely with the desires and preferences of individual users. This process requires users to articulate their ideas in words that are both comprehensible to the models and accur…

2021

RpBERT: A Text-image Relation Propagation-based BERT Model for Multimodal NER

AAAI 2021technical

Recently multimodal named entity recognition (MNER) has utilized images to improve the accuracy of NER in tweets. However, most of the multimodal methods use attention mechanisms to extract visual clues regardless of whether the text and image are relevant. Practically, the irrelevant text-image pai…

2020

RIVA: A Pre-trained Tweet Multimodal Model Based on Text-image Relation for Multimodal NER

COLING 2020main

Multimodal named entity recognition (MNER) for tweets has received increasing attention recently. Most of the multimodal methods used attention mechanisms to capture the text-related visual information. However, unrelated or weakly related text-image pairs account for a large proportion in tweets. V…

Cited by 35SourcePDFScholar