← Search

Yichu Zhou

4 accepted papers

2024

DocFormerv2: Local Features for Document Understanding

AAAI 2024technical

We propose DocFormerv2, a multi-modal transformer for Visual Document Understanding (VDU). The VDU domain entails understanding documents (beyond mere OCR predictions) e.g., extracting information from a form, VQA for documents and other tasks. VDU is challenging as it needs a model to make sense of…

Cited by 47SourcePDFScholar
2021

Putting Words in BERT’s Mouth: Navigating Contextualized Vector Spaces with Pseudowords

EMNLP 2021main

We present a method for exploring regions around individual points in a contextualized vector space (particularly, BERT space), as a way to investigate how these regions correspond to word senses. By inducing a contextualized “pseudoword” vector as a stand-in for a static embedding in the input laye…