2022
ScalableViT: Rethinking the Context-Oriented Generalization of Vision Transformer
ECCV 2022poster
"The vanilla self-attention mechanism inherently relies on pre-defined and steadfast computational dimensions. Such inflexibility restricts it from possessing context-oriented generalization that can bring more contextual cues and graphic representations. To mitigate this issue, we propose a Scalabl…