← Search

Duy-Kien Nguyen

3 accepted papers

2022

BoxeR: Box-Attention for 2D and 3D Transformers

CVPR 2022poster

In this paper, we propose a simple attention mechanism, we call Box-Attention. It enables spatial interaction between grid features, as sampled from boxes of interest, and improves the learning capability of transformers for several vision tasks. Specifically, we present BoxeR, short for Box Transfo…

Cited by 43PDFcodeScholar
2018

Improved Fusion of Visual and Language Representations by Dense Symmetric Co-Attention for Visual Question Answering

CVPR 2018poster

A key solution to visual question answering (VQA) exists in how to fuse visual and language features extracted from an input image and question. We show that an attention mechanism that enables dense, bi-directional interactions between the two modalities contributes to boost accuracy of prediction…

Cited by 379SourcePDFScholar