NAACL 2022long7 citations

Sketching as a Tool for Understanding and Accelerating Self-attention for Long Sequences

Yifan Chen, Qi Zeng, Dilek Hakkani-Tur, Di Jin, Heng Ji, Yun Yang

Abstract

Transformer-based models are not efficient in processing long sequences due to the quadratic space and time complexity of the self-attention modules. To address this limitation, Linformer and Informer reduce the quadratic complexity to linear (modulo logarithmic factors) via low-dimensional projection and row selection, respectively. These two models are intrinsically connected, and to understand their connection we introduce a theoretical framework of matrix sketching. Based on the theoretical analysis, we propose Skeinformer to accelerate self-attention and further improve the accuracy of matrix approximation to self-attention with column sampling, adaptive row normalization and pilot sampling reutilization. Experiments on the Long Range Arena benchmark demonstrate that our methods outperform alternatives with a consistently smaller time/space footprint.

BibTeX
@inproceedings{chen-etal-2022-sketching,
    title = "Sketching as a Tool for Understanding and Accelerating Self-attention for Long Sequences",
    author = "Chen, Yifan  and
      Zeng, Qi  and
      Hakkani-Tur, Dilek  and
      Jin, Di  and
      Ji, Heng  and
      Yang, Yun",
    editor = "Carpuat, Marine  and
      de Marneffe, Marie-Catherine  and
      Meza Ruiz, Ivan Vladimir",
    booktitle = "Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies",
    month = jul,
    year = "2022",
    address = "Seattle, United States",
    publisher = "Association for Computational Linguistics",
    url = "https://aclanthology.org/2022.naacl-main.381/",
    doi = "10.18653/v1/2022.naacl-main.381",
    pages = "5187--5199"
}
Sketching as a Tool for Understanding and Accelerating Self-attention for Long Sequences · NAACL 2022