← Search

Wenfeng Xie

9 accepted papers

2026

Appearance Discrepancy-guided Sequence Hybrid Masking for Robust Scene Text Recognition

AAAI 2026technical

Masked Image Modeling (MIM) has been widely recognized as a powerful self-supervised paradigm for learning general-purpose visual representations. However, standard MIM based on random masking tends to underperform in domain-specific tasks like Scene Text Recognition (STR), due to challenges such as

Cited by 0SourcePDFScholar
2024

Detection-Based Intermediate Supervision for Visual Question Answering

AAAI 2024technical

Recently, neural module networks (NMNs) have yielded ongoing success in answering compositional visual questions, especially those involving multi-hop visual and logical reasoning. NMNs decompose the complex question into several sub-tasks using instance-modules from the reasoning paths of that ques…

2024

Enhancing Low-Resource Relation Representations through Multi-View Decoupling

AAAI 2024technical

Recently, prompt-tuning with pre-trained language models (PLMs) has demonstrated the significantly enhancing ability of relation extraction (RE) tasks. However, in low-resource scenarios, where the available training data is scarce, previous prompt-based methods may still perform poorly for prompt-…

2024

Joint Multi-Facts Reasoning Network for Complex Temporal Question Answering Over Knowledge Graph

ICASSP 2024accepted

Temporal Knowledge Graph (TKG) is an extension of regular knowledge graph by attaching the time scope. Existing temporal knowledge graph question answering (TKGQA) models solely approach simple questions, owing to the prior assumption that each question only contains a single temporal fact with expl…

Cited by 0SourceScholar
2024

Mitigating Boundary Ambiguity and Inherent Bias for Text Classification in the Era of Large Language Models

ACL 2024findings

Text classification is a crucial task encountered frequently in practical scenarios, yet it is still under-explored in the era of large language models (LLMs). This study shows that LLMs are vulnerable to changes in the number and arrangement of options in text classification. Our extensive empirica…

2024

Position Debiasing Fine-Tuning for Causal Perception in Long-Term Dialogue

IJCAI 2024poster

The core of the dialogue system is to generate relevant, informative, and human-like responses based on extensive dialogue history. Recently, dialogue generation domain has seen mainstream adoption of large language models (LLMs), due to its powerful capability in generating utterances. However, the…

Cited by 2SourcePDFScholar
2024

Reinforcement Learning with Token-level Feedback for Controllable Text Generation

NAACL 2024findings

To meet the requirements of real-world applications, it is essential to control generations of large language models (LLMs). Prior research has tried to introduce reinforcement learning (RL) into controllable text generation while most existing methods suffer from overfitting issues (finetuning-base…

2023

TREA: Tree-Structure Reasoning Schema for Conversational Recommendation

ACL 2023long

Conversational recommender systems (CRS) aim to timely trace the dynamic interests of users through dialogues and generate relevant responses for item recommendations. Recently, various external knowledge bases (especially knowledge graphs) are incorporated into CRS to enhance the understanding of c…