2023
Building Vision Transformers with Hierarchy Aware Feature Aggregation
ICCV 2023poster
Thanks to the excellent global modeling capability of attention mechanisms, the Vision Transformer has achieved better results than ConvNet in many computer tasks. However, in generating hierarchical feature maps, the Transformer still adopts the ConvNet feature aggregation scheme. This leads to the…