Building an Icelandic Entity Linking Corpus

In this paper, we present the first Entity Linking corpus for Icelandic. We describe our approach of using a multilingual entity linking model (mGENRE) in combination with Wikipedia API Search (WAPIS) to label our data and compare it to an approach using WAPIS only. We find that our combined method reaches 53.9 coverage on our corpus, compared to 30.9 results and explain the value of using a multilingual system when working with Icelandic. Additionally, we analyze the data that remain unlabeled, identify patterns and discuss why they may be more difficult to annotate.

READ FULL TEXT

page 5

page 7

research
06/02/2021

MOLEMAN: Mention-Only Linking of Entities with a Mention Annotation Network

We present an instance-based nearest neighbor approach to entity linking...
research
12/22/2017

Ranking Triples using Entity Links in a Large Web Crawl - The Chicory Triple Scorer at WSDM Cup 2017

This paper describes the participation of team Chicory in the Triple Ran...
research
06/15/2023

Multilingual End to End Entity Linking

Entity Linking is one of the most common Natural Language Processing tas...
research
05/29/2018

Entity Linking in 40 Languages using MAG

A plethora of Entity Linking (EL) approaches has recently been developed...
research
05/15/2020

Neural Entity Linking on Technical Service Tickets

Entity linking, the task of mapping textual mentions to known entities, ...
research
07/17/2017

MAG: A Multilingual, Knowledge-base Agnostic and Deterministic Entity Linking Approach

Entity linking has recently been the subject of a significant body of re...
research
06/02/2020

REL: An Entity Linker Standing on the Shoulders of Giants

Entity linking is a standard component in modern retrieval system that i...

Please sign up or login with your details

Forgot password? Click here to reset