2022
Multi Resolution Analysis (MRA) for Approximate Self-Attention
ICML 2022spotlight
Transformers have emerged as a preferred model for many tasks in natural langugage processing and vision. Recent efforts on training and deploying Transformers more efficiently have identified many strategies to approximate the self-attention matrix, a key module in a Transformer architecture. Effec…