A Question-Answering Approach to Key Value Pair Extraction from Form-like Document Images

04/17/2023
by   Kai Hu, et al.
0

In this paper, we present a new question-answering (QA) based key-value pair extraction approach, called KVPFormer, to robustly extracting key-value relationships between entities from form-like document images. Specifically, KVPFormer first identifies key entities from all entities in an image with a Transformer encoder, then takes these key entities as questions and feeds them into a Transformer decoder to predict their corresponding answers (i.e., value entities) in parallel. To achieve higher answer prediction accuracy, we propose a coarse-to-fine answer prediction approach further, which first extracts multiple answer candidates for each identified question in the coarse stage and then selects the most likely one among these candidates in the fine stage. In this way, the learning difficulty of answer prediction can be effectively reduced so that the prediction accuracy can be improved. Moreover, we introduce a spatial compatibility attention bias into the self-attention/cross-attention mechanism for to better model the spatial interactions between entities. With these new techniques, our proposed achieves state-of-the-art results on FUNSD and XFUND datasets, outperforming the previous best-performing method by 7.2% and 13.2% in F1 score, respectively.

READ FULL TEXT
research
01/03/2019

Coarse-grain Fine-grain Coattention Network for Multi-evidence Question Answering

End-to-end neural models have made significant progress in question answ...
research
04/03/2018

Improved Fusion of Visual and Language Representations by Dense Symmetric Co-Attention for Visual Question Answering

A key solution to visual question answering (VQA) exists in how to fuse ...
research
01/16/2022

Double Retrieval and Ranking for Accurate Question Answering

Recent work has shown that an answer verification step introduced in Tra...
research
03/10/2015

A Case Based Reasoning Approach for Answer Reranking in Question Answering

In this document I present an approach to answer validation and rerankin...
research
05/02/2019

Conditioning LSTM Decoder and Bi-directional Attention Based Question Answering System

Applying neural-networks on Question Answering has gained increasing pop...
research
10/19/2020

Knowledge-guided Open Attribute Value Extraction with Reinforcement Learning

Open attribute value extraction for emerging entities is an important bu...
research
12/15/2021

Responsive parallelized architecture for deploying deep learning models in production environments

Recruiters can easily shortlist candidates for jobs via viewing their cu...

Please sign up or login with your details

Forgot password? Click here to reset