Corpus-level Fine-grained Entity Typing

08/07/2017
by   Yadollah Yaghoobzadeh, et al.
0

This paper addresses the problem of corpus-level entity typing, i.e., inferring from a large corpus that an entity is a member of a class such as "food" or "artist". The application of entity typing we are interested in is knowledge base completion, specifically, to learn which classes an entity is a member of. We propose FIGMENT to tackle this problem. FIGMENT is embedding- based and combines (i) a global model that scores based on aggregated contextual information of an entity and (ii) a context model that first scores the individual occurrences of an entity and then aggregates the scores. Each of the two proposed models has some specific properties. For the global model, learning high quality entity representations is crucial because it is the only source used for the predictions. Therefore, we introduce representations using name and contexts of entities on the three levels of entity, word, and character. We show each has complementary information and a multi-level representation is the best. For the context model, we need to use distant supervision since the context-level labels are not available for entities. Distant supervised labels are noisy and this harms the performance of models. Therefore, we introduce and apply new algorithms for noise mitigation using multi-instance learning. We show the effectiveness of our models in a large entity typing dataset, built from Freebase.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/25/2016

Corpus-level Fine-grained Entity Typing Using Contextual Information

This paper addresses the problem of corpus-level entity typing, i.e., in...
research
01/08/2017

Multi-level Representations for Fine-Grained Typing of Knowledge Base Entities

Entities are essential elements of natural language. In this paper, we p...
research
04/07/2020

Fine-Grained Named Entity Typing over Distantly Supervised Data Based on Refined Representations

Fine-Grained Named Entity Typing (FG-NET) is a key component in Natural ...
research
12/22/2016

Noise Mitigation for Neural Entity Typing and Relation Extraction

In this paper, we address two different types of noise in information ex...
research
05/17/2019

Distant Learning for Entity Linking with Automatic Noise Detection

Accurate entity linkers have been produced for domains and languages whe...
research
04/26/2018

Open Information Extraction with Global Structure Constraints

Extracting entities and their relations from text is an important task f...
research
12/13/2018

Same but Different: Distant Supervision for Predicting and Understanding Entity Linking Difficulty

Entity Linking (EL) is the task of automatically identifying entity ment...

Please sign up or login with your details

Forgot password? Click here to reset