2023
Cross-Modal Orthogonal High-Rank Augmentation for RGB-Event Transformer-Trackers
ICCV 2023poster
This paper addresses the problem of cross-modal object tracking from RGB videos and event data. Rather than constructing a complex cross-modal fusion network, we explore the great potential of a pre-trained vision Transformer (ViT). Particularly, we delicately investigate plug-and-play training augm…