← Search

Zhuoyao Zhong

3 accepted papers

2023

A Question-Answering Approach to Key Value Pair Extraction from Form-Like Document Images

AAAI 2023technical

In this paper, we present a new question-answering (QA) based key-value pair extraction approach, called KVPFormer, to robustly extracting key-value relationships between entities from form-like document images. Specifically, KVPFormer first identifies key entities from all entities in an image with…

Cited by 13SourcePDFScholar
2023

Exploring Predicate Visual Context in Detecting of Human-Object Interactions

ICCV 2023poster

Recently, the DETR framework has emerged as the dominant approach for human--object interaction (HOI) research. In particular, two-stage transformer-based HOI detectors are amongst the most performant and training-efficient approaches. However, these often condition HOI classification on object feat…

Cited by 50PDFcodeScholar
2017

DeepText: A new approach for text proposal generation and text detection in natural images

ICASSP 2017accepted

In this paper, we develop a new approach called DeepText for text region proposal generation and text detection in natural images via a fully convolutional neural network (CNN). First, we propose the novel inception region proposal network (Inception-RPN), which slides an inception network with mult…

Cited by 0SourceScholar