A Novel Framework to Expedite Systematic Reviews by Automatically Building Information Extraction Training Corpora

06/21/2016
by   Tanmay Basu, et al.
0

A systematic review identifies and collates various clinical studies and compares data elements and results in order to provide an evidence based answer for a particular clinical question. The process is manual and involves lot of time. A tool to automate this process is lacking. The aim of this work is to develop a framework using natural language processing and machine learning to build information extraction algorithms to identify data elements in a new primary publication, without having to go through the expensive task of manual annotation to build gold standards for each data element type. The system is developed in two stages. Initially, it uses information contained in existing systematic reviews to identify the sentences from the PDF files of the included references that contain specific data elements of interest using a modified Jaccard similarity measure. These sentences have been treated as labeled data.A Support Vector Machine (SVM) classifier is trained on this labeled data to extract data elements of interests from a new article. We conducted experiments on Cochrane Database systematic reviews related to congestive heart failure using inclusion criteria as an example data element. The empirical results show that the proposed system automatically identifies sentences containing the data element of interest with a high recall (93.75 (27.05 average). The empirical results suggest that the tool is retrieving valuable information from the reference articles, even when it is time-consuming to identify them manually. Thus we hope that the tool will be useful for automatic data extraction from biomedical research publications. The future scope of this work is to generalize this information framework for all types of systematic reviews.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
09/20/2017

A shared latent space matrix factorisation method for recommending new trial evidence for systematic review updates

Clinical trial registries can be used to monitor the production of trial...
research
01/29/2018

Improving Active Learning in Systematic Reviews

Systematic reviews are essential to summarizing the results of different...
research
10/09/2020

Scaling Systematic Literature Reviews with Machine Learning Pipelines

Systematic reviews, which entail the extraction of data from large numbe...
research
11/03/2018

Unsupervised Identification of Study Descriptors in Toxicology Research: An Experimental Study

Identifying and extracting data elements such as study descriptors in pu...
research
07/12/2021

A Systematic Literature Review of Automated ICD Coding and Classification Systems using Discharge Summaries

Codification of free-text clinical narratives have long been recognised ...
research
08/11/2018

The Impact of Automatic Pre-annotation in Clinical Note Data Element Extraction - the CLEAN Tool

Objective. Annotation is expensive but essential for clinical note revie...
research
02/12/2021

A Visual Analysis Approach to Update Systematic Reviews

Context: In order to preserve the value of Systematic Reviews (SRs), the...

Please sign up or login with your details

Forgot password? Click here to reset