← Search

Yunling Zheng

2 accepted papers

2026

SEMA: a Scalable and Efficient Mamba like Attention via Token Localization and Averaging

ICML 2026poster

Attention is the critical component of a transformer. Yet the quadratic computational complexity of vanilla full attention in the input size and the inability of its linear attention variant to focus have been challenges for computer vision tasks. We provide a mathematical definition of generalized …

Cited by 0SourceScholar
2022

Glassoformer: A Query-Sparse Transformer for Post-Fault Power Grid Voltage Prediction

ICASSP 2022accepted

We propose GLassoformer, a novel and efficient transformer architecture leveraging group Lasso regularization to reduce the number of queries of the standard self-attention mechanism. Due to the sparsified queries, GLassoformer is more computationally efficient than the standard transformers. On the…

Cited by 0SourceScholar