arXiv · 2010.05322
Revising FUNSD dataset for key-value detection in document images
Abstract
FUNSD is one of the limited publicly available datasets for information extraction from document im-ages. The information in the FUNSD dataset is defined by text areas of four categories ("key", "value", "header", "other", and "background") and connectivity between areas as key-value relations. In-specting FUNSD, we found several inconsistency in labeling, which impeded its applicability to thekey-value extraction problem. In this report, we described some labeling issues in FUNSD and therevision we made to the dataset. We also reported our implementation of for key-value detection onFUNSD using a UNet model as baseline results and an improved UNet model with Channel-InvariantDeformable Convolution.
Explore related subjects
Keep this discovery
Hieu M. Vu, Diep Thi-Ngoc Nguyen. 2020-10-11. Revising FUNSD dataset for key-value detection in document images. https://arxiv.org/abs/2010.05322
Cite the original work for its findings. Save a collection to share your selection of sources.