2017
Learning to Extract Semantic Structure From Documents Using Multimodal Fully Convolutional Neural Networks
CVPR 2017spotlight
We present an end-to-end, multimodal, fully convolutional network for extracting semantic structures from document images. We consider document semantic structure extraction as a pixel-wise segmentation task, and propose a unified model that classifies pixels based not only on their visual appearanc…