← Search

Hamid Reza Vaezi Joze

3 accepted papers

2022

Adaptive Token Sampling for Efficient Vision Transformers

ECCV 2022poster

"While state-of-the-art vision transformer models achieve promising results in image classification, they are computationally expensive and require many GFLOPs. Although the GFLOPs of a vision transformer can be decreased by reducing the number of tokens in the network, there is no setting that is o…

2020

MMTM: Multimodal Transfer Module for CNN Fusion

CVPR 2020poster

In late fusion, each modality is processed in a separate unimodal Convolutional Neural Network (CNN) stream and the scores of each modality are fused at the end. Due to its simplicity, late fusion is still the predominant approach in many state-of-the-art multimodal applications. In this paper, we p…

Cited by 397PDFScholar
2019

Improving the Performance of Unimodal Dynamic Hand-Gesture Recognition With Multimodal Training

CVPR 2019poster

We present an efficient approach for leveraging the knowledge from multiple modalities in training unimodal 3D convolutional neural networks (3D-CNNs) for the task of dynamic hand gesture recognition. Instead of explicitly combining multimodal information, which is commonplace in many state-of-the…

Cited by 212PDFScholar