2026
LaplacianFormer:Rethinking Linear Attention with Laplacian Kernel
ICLR 2026poster
The quadratic complexity of softmax attention presents a major obstacle for scaling Transformers to high-resolution vision tasks. Existing linear attention variants often replace the softmax with Gaussian kernels to reduce complexity, but such approximations lack theoretical grounding and tend to ov…