← Search

Xiaofeng Ren

3 accepted papers

2022

Doubly-Fused ViT: Fuse Information from Vision Transformer Doubly with Local Representation

ECCV 2022poster

"Vision Transformer (ViT) has recently emerged as a new paradigm for computer vision tasks, but is not as efficient as convolutional neural networks (CNN). In this paper, we propose an efficient ViT architecture, named Doubly-Fused ViT (DFvT), where we feed low-resolution feature maps to self-attent…

2022

SCMT: Self-Correction Mean Teacher for Semi-supervised Object Detection

IJCAI 2022poster

Semi-Supervised Object Detection (SSOD) aims to improve performance by leveraging a large amount of unlabeled data. Existing works usually adopt the teacher-student framework to enforce student to learn consistent predictions over the pseudo-labels generated by teacher. However, the performance of t…

Cited by 8SourcePDFScholar