← Search

Fu Lee Wang

14 accepted papers

2026

Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence

IJCAI 2026

Reliable reasoning in Large Language Models (LLMs) is challenged by their propensity for hallucination. While augmenting LLMs with Knowledge Graphs (KGs) improves factual accuracy, existing KG-augmented methods fail to quantify epistemic uncertainty in both the retrieved evidence and LLMs' reasoning

Cited by 0Scholar
2026

RGGT: A Generative-Prior-Guided Transformer for Unified Rigid and Non-Rigid Point Cloud Registration

ICML 2026poster

Point cloud registration can be categorized into rigid and non-rigid settings depending on the motion characteristics of the underlying objects. Rigid alignment assumes a single global transformation under which corresponding points remain geometrically consistent across scales, whereas non-rigid al…

Cited by 0SourceScholar
2025

DeblurDiff: Real-Word Image Deblurring with Generative Diffusion Models

NeurIPS 2025poster

Diffusion models have achieved significant progress in image generation and the pre-trained Stable Diffusion (SD) models are helpful for image deblurring by providing clear image priors. However, directly using a blurry image or a pre-deblurred one as a conditional control for SD will either hinder…

Cited by 0SourceScholar
2025

PairEdit: Learning Semantic Variations for Exemplar-based Image Editing

NeurIPS 2025poster

Recent advancements in text-guided image editing have achieved notable success by leveraging natural language prompts for fine-grained semantic control. However, certain editing semantics are challenging to specify precisely using textual descriptions alone. A practical alternative involves learning…

Cited by 0SourcecodeScholar
2025

Span Attention for Entity-Consistent Task-Oriented Dialogue Response Generation

ICASSP 2025accepted

Task-oriented dialogue systems have recently gained increasing attention due to their capability of using natural language to fulfill specific user demands, such as restaurant reservation and hotel booking. Recent works directly model task-oriented dialogue response as a text generation task. Howeve…

Cited by 0SourceScholar
2025

When Allies Turn Foes: Exploring Group Characteristics of LLM-Based Multi-Agent Collaborative Systems Under Adversarial Attacks

EMNLP 2025

This paper investigates the group characteristics in multi-agent collaborative systems under adversarial attacks. Adversarial agents are tasked with generating counterfactual answers to a given collaborative problem, while collaborative agents normally interact with other agents to solve the given p

2024

AttnDreamBooth: Towards Text-Aligned Personalized Text-to-Image Generation

NeurIPS 2024poster

Recent advances in text-to-image models have enabled high-quality personalized image synthesis based on user-provided concepts with flexible textual control. In this work, we analyze the limitations of two primary techniques in text-to-image personalization: Textual Inversion and DreamBooth. When in…

Cited by 5SourcePDFScholar
2024

MTA: A Lightweight Multilingual Text Alignment Model for Cross-Language Visual Word Sense Disambiguation

ICASSP 2024accepted

Visual Word Sense Disambiguation (Visual-WSD), as a sub-task of fine-grained image-text retrieval, requires a high level of language-vision understanding to capture and exploit the nuanced relationships between text and visual features. However, the cross-linguistic background only with limited cont…

Cited by 0SourceScholar
2024

PolCLIP: A Unified Image-Text Word Sense Disambiguation Model via Generating Multimodal Complementary Representations

ACL 2024long

Word sense disambiguation (WSD) can be viewed as two subtasks: textual word sense disambiguation (Textual-WSD) and visual word sense disambiguation (Visual-WSD). They aim to identify the most semantically relevant senses or images to a given context containing ambiguous target words. However, existi…

2023

Geogcn: Geometric Dual-Domain Graph Convolution Network For Point Cloud Denoising

ICASSP 2023accepted

We propose GeoGCN, a novel geometric dual-domain graph convolution network for point cloud denoising (PCD). Beyond the traditional wisdom of PCD, to fully exploit the geometric information of point clouds, we define two kinds of surface normals, one is called Real Normal (RN), and the other is Virtu…

Cited by 0SourceScholar
2023

Recurrent Attention Networks for Long-text Modeling

ACL 2023findings

Self-attention-based models have achieved remarkable progress in short-text mining. However, the quadratic computational complexities restrict their application in long text processing. Prior works have adopted the chunking strategy to divide long documents into chunks and stack a self-attention bac…

2022

A Self-supervised Joint Training Framework for Document Reranking

NAACL 2022findings

Pretrained language models such as BERT have been successfully applied to a wide range of natural language processing tasks and also achieved impressive performance in document reranking tasks. Recent works indicate that further pretraining the language models on the task-specific datasets before fi…

Cited by 2SourcePDFScholar
2020

Detail-recovery Image Deraining via Context Aggregation Networks

CVPR 2020poster

This paper looks at this intriguing question: are single images with their details lost during deraining, reversible to their artifact-free status? We propose an end-to-end detail-recovery image deraining network (termed a DRDNet) to solve the problem. Unlike existing image deraining approaches that…

Cited by 217PDFcodeScholar
2020

Geometry and Learning Co-Supported Normal Estimation for Unstructured Point Cloud

CVPR 2020poster

In this paper, we propose a normal estimation method for unstructured point cloud. We observe that geometric estimators commonly focus more on feature preservation but are hard to tune parameters and sensitive to noise, while learning-based approaches pursue an overall normal estimation accuracy but…

Cited by 42PDFScholar