← Search

Zhang Xiong

10 accepted papers

2025

NoiseHGNN: Synthesized Similarity Graph-Based Neural Network for Noised Heterogeneous Graph Representation Learning

AAAI 2025technical

Real-world graph data environments intrinsically exist noise (e.g., link and structure errors) that inevitably disturb the effectiveness of graph representation and downstream learning tasks. For homogeneous graphs, the latest works use original node features to synthesize a similarity graph that ca…

2023

Token-Level Self-Evolution Training for Sequence-to-Sequence Learning

ACL 2023short

Adaptive training approaches, widely used in sequence-to-sequence models, commonly reweigh the losses of different target tokens based on priors, e.g. word frequency. However, most of them do not consider the variation of learning difficulty in different training steps, and overly emphasize the lear…

Cited by 23SourcePDFScholar
2023

Transformer-Patcher: One Mistake Worth One Neuron

ICLR 2023poster

Large Transformer-based Pretrained Language Models (PLMs) dominate almost all Natural Language Processing (NLP) tasks. Nevertheless, they still make mistakes from time to time. For a model deployed in an industrial environment, fixing these mistakes quickly and robustly is vital to improve user expe…

2022

Mixture of Attention Heads: Selecting Attention Heads Per Token

EMNLP 2022main

Mixture-of-Experts (MoE) networks have been proposed as an efficient way to scale up model capacity and implement conditional computing. However, the study of MoE components mostly focused on the feedforward layer in Transformer architecture. This paper proposes the Mixture of Attention Heads (MoA),…

2021

Paragraph Level Multi-Perspective Context Modeling for Question Generation

ICASSP 2021accepted

Proper understanding of paragraph is essential for question generation task since the semantic interaction is complicated among sentences. How to integrate long text paragraph information into question generation is still a challenge. In this research, we proposed a multi-perspective paragraph conte…

Cited by 0SourceScholar
2021

Topic-Aware Dialogue Generation with Two-Hop Based Graph Attention

ICASSP 2021accepted

Generating on-topic responses and understanding the background information of context are both significant for dialogue generation. However, few works simultaneously concentrate on these two issues. For this purpose, we propose an open-domain topic-aware dialogue generation model via joint learning.…

Cited by 0SourceScholar
2020

DR-KFS: A Differentiable Visual Similarity Metric for 3D Shape Reconstruction

ECCV 2020poster

We introduce a differential visual similarity metric to train deep neural networks for 3D reconstruction, aimed at improving reconstruction quality. The metric compares two 3D shapes by measuring distances between multi-view images differentiably rendered from the shapes. Importantly, the image-spac…

Cited by 9SourcePDFScholar
2015

Confidence Preserving Machine for Facial Action Unit Detection

ICCV 2015poster

Varied sources of error contribute to the challenge of facial action unit detection. Previous approaches address specific and known sources. However, many sources are unknown. To address the ubiquity of error, we propose a Confident Preserving Machine (CPM) that follows an easy-to-hard classificatio…

Cited by 82PDFScholar