← Search

Sujoy Paul*

1 accepted papers

2024

LookupViT: Compressing visual information to a limited number of tokens

ECCV 2024poster

"Vision Transformers (ViT) have emerged as the de-facto choice for numerous industry grade vision solutions. But their inference cost can be prohibitive for many settings, as they compute self-attention in each layer which suffers from quadratic computational complexity in the number of tokens. On t…

Cited by 11SourcePDFScholar