Retrievability in an Integrated Retrieval System: An Extended Study

03/27/2023
by   Dwaipayan Roy, et al.
0

Retrievability measures the influence a retrieval system has on the access to information in a given collection of items. This measure can help in making an evaluation of the search system based on which insights can be drawn. In this paper, we investigate the retrievability in an integrated search system consisting of items from various categories, particularly focussing on datasets, publications and variables in a real-life Digital Library (DL). The traditional metrics, that is, the Lorenz curve and Gini coefficient, are employed to visualize the diversity in retrievability scores of the three retrievable document types (specifically datasets, publications, and variables). Our results show a significant popularity bias with certain items being retrieved more often than others. Particularly, it has been shown that certain datasets are more likely to be retrieved than other datasets in the same category. In contrast, the retrievability scores of items from the variable or publication category are more evenly distributed. We have observed that the distribution of document retrievability is more diverse for datasets as compared to publications and variables.

READ FULL TEXT
research
05/02/2022

Studying Retrievability of Publications and Datasets in an Integrated Retrieval System

In this paper, we investigate the retrievability of datasets and publica...
research
06/04/2020

Characteristics of Dataset Retrieval Sessions: Experiences from a Real-life Digital Library

Secondary analysis or the reuse of existing survey data is a common prac...
research
03/16/2017

The coverage of Microsoft Academic: Analyzing the publication output of a university

This is the first detailed study on the coverage of Microsoft Academic (...
research
05/28/2019

The HyperBagGraph DataEdron: An Enriched Browsing Experience of Multimedia Datasets

Traditional verbatim browsers give back information in a linear way acco...
research
09/13/2021

An Adaptive Boosting Technique to Mitigate Popularity Bias in Recommender System

The observed ratings in most recommender systems are subjected to popula...
research
12/12/2022

Multivariate Powered Dirichlet Hawkes Process

The publication time of a document carries a relevant information about ...

Please sign up or login with your details

Forgot password? Click here to reset