← Search

Youbing Hu

1 accepted papers

2024

LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition

AAAI 2024technical

The Vision Transformer (ViT) excels in accuracy when handling high-resolution images, yet it confronts the challenge of significant spatial redundancy, leading to increased computational and memory requirements. To address this, we present the Localization and Focus Vision Transformer (LF-ViT). This…