Cardinal Graph Convolution Framework for Document Information Extraction

7Citations
Citations of this article
16Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Graph Convolutional Networks (GCN) have been recognized as successful for processing pseudo-spatial graph representations of the underlying structure of documents. We present Cardinal Graph Convolutional Networks (CGCN), an efficient and flexible extension of GCNs with cardinal-direction awareness of spatial node arrangement. The formulation of CGCNs retains the traditional GCN permutation invariance, ensuring directional neighbors are involved in learning abstract representations, even in the absence of a proper ordering of the nodes. We show that CGCNs achieve state of the art results on an invoice information extraction task, jointly learning a word-level tagging as well as document meta-level classification and regression. We also present a new multiscale Inception-like CGCN block-layer, as well as Conv-Pool-DeConv-DePool UNet-like architecture, which increase the receptive field. We demonstrate the utility of CGCNs on private and public datasets, with respect to several baseline models: sequential LSTM, transformer classifier, non-cardinal GCNs, and an image-convolutional approach.

Cite

CITATION STYLE

APA

Gal, R., Ardazi, S., & Shilkrot, R. (2020). Cardinal Graph Convolution Framework for Document Information Extraction. In Proceedings of the ACM Symposium on Document Engineering, DocEng 2020. Association for Computing Machinery, Inc. https://doi.org/10.1145/3395027.3419584

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free