← Search

Bhargava Urala Kota

1 accepted papers

2021

DocFormer: End-to-End Transformer for Document Understanding

ICCV 2021poster

We present DocFormer - a multi-modal transformer based architecture for the task of Visual Document Understanding (VDU). VDU is a challenging problem which aims to understand documents in their varied formats(forms, receipts etc.) and layouts. In addition, DocFormer is pre-trained in an unsupervised…

Cited by 342PDFScholar