← Search

Dehong Ma

3 accepted papers

2024

UEGP: Unified Expert-Guided Pre-training for Knowledge Rekindle

NAACL 2024findings

Pre-training and fine-tuning framework has become the standard training paradigm for NLP tasks and is also widely used in industrial-level applications. However, there are still a limitation with this paradigm: simply fine-tuning with task-specific objectives tends to converge to local minima, resul…

2022

A Question-Oriented Propagation Network for News Reading Comprehension

ICASSP 2022accepted

Machine reading comprehension of news articles remains to be a challenging task since the lengths of its context documents are long. Such reading comprehension task usually requires document-level language understanding while state-of-the-art, pretrained question answering models can only encode seq…

Cited by 0SourceScholar
2022

PILE: Pairwise Iterative Logits Ensemble for Multi-Teacher Labeled Distillation

EMNLP 2022industry

Pre-trained language models have become a crucial part of ranking systems and achieved very impressive effects recently. To maintain high performance while keeping efficient computations, knowledge distillation is widely used. In this paper, we focus on two key questions in knowledge distillation fo…

Cited by 4SourcePDFScholar