A Multilingual Entity Linking System for Wikipedia with a Machine-in-the-Loop Approach

05/31/2021
by   Martin Gerlach, et al.
0

Hyperlinks constitute the backbone of the Web; they enable user navigation, information discovery, content ranking, and many other crucial services on the Internet. In particular, hyperlinks found within Wikipedia allow the readers to navigate from one page to another to expand their knowledge on a given subject of interest or to discover a new one. However, despite Wikipedia editors' efforts to add and maintain its content, the distribution of links remains sparse in many language editions. This paper introduces a machine-in-the-loop entity linking system that can comply with community guidelines for adding a link and aims at increasing link coverage in new pages and wiki-projects with low-resources. To tackle these challenges, we build a context and language agnostic entity linking model that combines data collected from millions of anchors found across wiki-projects, as well as billions of users' reading sessions. We develop an interactive recommendation interface that proposes candidate links to editors who can confirm, reject, or adapt the recommendation with the overall aim of providing a more accessible editing experience for newcomers through structured tasks. Our system's design choices were made in collaboration with members of several language communities. When the system is implemented as part of Wikipedia, its usage by volunteer editors will help us build a continuous evaluation dataset with active feedback. Our experimental results show that our link recommender can achieve a precision above 80 ensuring a recall of at least 50 continents, and families.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
05/02/2017

Entity Linking with people entity on Wikipedia

This paper introduces a new model that uses named entity recognition, co...
research
03/13/2019

Overview of the Ugglan Entity Discovery and Linking System

Ugglan is a system designed to discover named entities and link them to ...
research
05/28/2020

Empirical Evaluation of Pretraining Strategies for Supervised Entity Linking

In this work, we present an entity linking model which combines a Transf...
research
05/03/2018

Deep Linking Desktop Resources

Deep Linking is the process of referring to a specific piece of web cont...
research
05/25/2021

Predicting Links on Wikipedia with Anchor Text Information

Wikipedia, the largest open-collaborative online encyclopedia, is a corp...
research
04/15/2020

Layered Graph Embedding for Entity Recommendation using Wikipedia in the Yahoo! Knowledge Graph

In this paper, we describe an embedding-based entity recommendation fram...
research
07/16/2020

Wikipedia's Network Bias on Controversial Topics

The most important feature of Wikipedia is the presence of hyperlinks in...

Please sign up or login with your details

Forgot password? Click here to reset