Logician: A Unified End-to-End Neural Approach for Open-Domain Information Extraction

04/29/2019
by   Mingming Sun, et al.
6

In this paper, we consider the problem of open information extraction (OIE) for extracting entity and relation level intermediate structures from sentences in open-domain. We focus on four types of valuable intermediate structures (Relation, Attribute, Description, and Concept), and propose a unified knowledge expression form, SAOKE, to express them. We publicly release a data set which contains more than forty thousand sentences and the corresponding facts in the SAOKE format labeled by crowd-sourcing. To our knowledge, this is the largest publicly available human labeled data set for open information extraction tasks. Using this labeled SAOKE data set, we train an end-to-end neural model using the sequenceto-sequence paradigm, called Logician, to transform sentences into facts. For each sentence, different to existing algorithms which generally focus on extracting each single fact without concerning other possible facts, Logician performs a global optimization over all possible involved facts, in which facts not only compete with each other to attract the attention of words, but also cooperate to share words. An experimental study on various types of open domain relation extraction tasks reveals the consistent superiority of Logician to other states-of-the-art algorithms. The experiments verify the reasonableness of SAOKE format, the valuableness of SAOKE data set, the effectiveness of the proposed Logician model, and the feasibility of the methodology to apply end-to-end learning paradigm on supervised data sets for the challenging tasks of open information extraction.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/26/2018

Open Information Extraction with Global Structure Constraints

Extracting entities and their relations from text is an important task f...
research
07/24/2017

Analysing Errors of Open Information Extraction Systems

We report results on benchmarking Open Information Extraction (OIE) syst...
research
09/15/2021

AnnIE: An Annotation Platform for Constructing Complete Open Information Extraction Benchmark

Open Information Extraction (OIE) is the task of extracting facts from s...
research
04/26/2018

Integrating Local Context and Global Cohesiveness for Open Information Extraction

Extracting entities and their relations from text is an important task f...
research
05/07/2023

Shall We Trust All Relational Tuples by Open Information Extraction? A Study on Speculation Detection

Open Information Extraction (OIE) aims to extract factual relational tup...
research
08/21/2018

Neural Relation Extraction via Inner-Sentence Noise Reduction and Transfer Learning

Extracting relations is critical for knowledge base completion and const...
research
12/17/2020

InSRL: A Multi-view Learning Framework Fusing Multiple Information Sources for Distantly-supervised Relation Extraction

Distant supervision makes it possible to automatically label bags of sen...

Please sign up or login with your details

Forgot password? Click here to reset