← Search

Shengqiang Zhang

1 accepted papers

2022

Attention Temperature Matters in Abstractive Summarization Distillation

ACL 2022long

Recent progress of abstractive text summarization largely relies on large pre-trained sequence-to-sequence Transformer models, which are computationally expensive. This paper aims to distill these large models into smaller ones for faster inference and with minimal performance loss. Pseudo-labeling…