CoType: Joint Extraction of Typed Entities and Relations with Knowledge Bases

10/27/2016
by   Xiang Ren, et al.
0

Extracting entities and relations for types of interest from text is important for understanding massive text corpora. Traditionally, systems of entity relation extraction have relied on human-annotated corpora for training and adopted an incremental pipeline. Such systems require additional human expertise to be ported to a new domain, and are vulnerable to errors cascading down the pipeline. In this paper, we investigate joint extraction of typed entities and relations with labeled data heuristically obtained from knowledge bases (i.e., distant supervision). As our algorithm for type labeling via distant supervision is context-agnostic, noisy training data poses unique challenges for the task. We propose a novel domain-independent framework, called CoType, that runs a data-driven text segmentation algorithm to extract entity mentions, and jointly embeds entity mentions, relation mentions, text features and type labels into two low-dimensional spaces (for entity and relation mentions respectively), where, in each space, objects whose types are close will also have similar representations. CoType, then using these learned embeddings, estimates the types of test (unlinkable) mentions. We formulate a joint optimization problem to learn embeddings from text corpora and knowledge bases, adopting a novel partial-label loss function for noisy labeled data and introducing an object "translation" function to capture the cross-constraints of entities and relations on each other. Experiments on three public datasets demonstrate the effectiveness of CoType across different domains (e.g., news, biomedical), with an average of 25 next best method.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
02/17/2016

Label Noise Reduction in Entity Typing by Heterogeneous Partial-Label Embedding

Current systems of fine-grained entity typing use distant supervision in...
research
10/30/2017

Indirect Supervision for Relation Extraction using Question-Answer Pairs

Automatic relation extraction (RE) for types of interest is of great imp...
research
06/26/2019

Eliciting Knowledge from Experts:Automatic Transcript Parsing for Cognitive Task Analysis

Cognitive task analysis (CTA) is a type of analysis in applied psycholog...
research
04/26/2018

Integrating Local Context and Global Cohesiveness for Open Information Extraction

Extracting entities and their relations from text is an important task f...
research
04/26/2018

Open Information Extraction with Global Structure Constraints

Extracting entities and their relations from text is an important task f...
research
06/25/2017

Automatic Synonym Discovery with Knowledge Bases

Recognizing entity synonyms from text has become a crucial task in many ...
research
04/14/2017

Cardinal Virtues: Extracting Relation Cardinalities from Text

Information extraction (IE) from text has largely focused on relations b...

Please sign up or login with your details

Forgot password? Click here to reset