← Search

Srinivasan Sivanandan

1 accepted papers

2024

Channel Vision Transformers: An Image Is Worth 1 x 16 x 16 Words

ICLR 2024poster

Vision Transformer (ViT) has emerged as a powerful architecture in the realm of modern computer vision. However, its application in certain imaging fields, such as microscopy and satellite imaging, presents unique challenges. In these domains, images often contain multiple channels, each carrying se…