A Self-supervised Approach for Semantic Indexing in the Context of COVID-19 Pandemic

10/07/2020
by   Nima Ebadi, et al.
0

The pandemic has accelerated the pace at which COVID-19 scientific papers are published. In addition, the process of manually assigning semantic indexes to these papers by experts is even more time-consuming and overwhelming in the current health crisis. Therefore, there is an urgent need for automatic semantic indexing models which can effectively scale-up to newly introduced concepts and rapidly evolving distributions of the hyperfocused related literature. In this research, we present a novel semantic indexing approach based on the state-of-the-art self-supervised representation learning and transformer encoding exclusively suitable for pandemic crises. We present a case study on a novel dataset that is based on COVID-19 papers published and manually indexed in PubMed. Our study shows that our self-supervised model outperforms the best performing models of BioASQ Task 8a by micro-F1 score of 0.1 and LCA-F score of 0.08 on average. Our model also shows superior performance on detecting the supplementary concepts which is quite important when the focus of the literature has drastically shifted towards specific concepts related to the pandemic. Our study sheds light on the main challenges confronting semantic indexing models during a pandemic, namely new domains and drastic changes of their distributions, and as a superior alternative for such situations, propose a model founded on approaches which have shown auspicious performance in improving generalization and data efficiency in various NLP tasks. We also show the joint indexing of major Medical Subject Headings (MeSH) and supplementary concepts improves the overall performance.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/15/2021

Interpretable Self-supervised Multi-task Learning for COVID-19 Information Retrieval and Extraction

The rapidly evolving literature of COVID-19 related articles makes it ch...
research
05/13/2020

MeSH descriptors indicate the knowledge growth in the SARS-CoV-2/COVID-19 pandemic

The scientific papers dealing with the novel betacoronavirus SARS-CoV-2 ...
research
05/13/2020

Meta-Research: COVID-19 medical papers have fewer women first authors than expected

The COVID-19 pandemic has resulted in school closures and distancing req...
research
08/03/2022

Multi-Feature Vision Transformer via Self-Supervised Representation Learning for Improvement of COVID-19 Diagnosis

The role of chest X-ray (CXR) imaging, due to being more cost-effective,...
research
01/09/2022

Zero-Shot and Few-Shot Classification of Biomedical Articles in Context of the COVID-19 Pandemic

MeSH (Medical Subject Headings) is a large thesaurus created by the Nati...
research
10/17/2021

Prioritization of COVID-19-related literature via unsupervised keyphrase extraction and document representation learning

The COVID-19 pandemic triggered a wave of novel scientific literature th...

Please sign up or login with your details

Forgot password? Click here to reset