Visual Exploration and Knowledge Discovery from Biomedical Dark Data

09/28/2020
by   Shashwat Aggarwal, et al.
0

Data visualization techniques proffer efficient means to organize and present data in graphically appealing formats, which not only speeds up the process of decision making and pattern recognition but also enables decision-makers to fully understand data insights and make informed decisions. Over time, with the rise in technological and computational resources, there has been an exponential increase in the world's scientific knowledge. However, most of it lacks structure and cannot be easily categorized and imported into regular databases. This type of data is often termed as Dark Data. Data visualization techniques provide a promising solution to explore such data by allowing quick comprehension of information, the discovery of emerging trends, identification of relationships and patterns, etc. In this empirical research study, we use the rich corpus of PubMed comprising of more than 30 million citations from biomedical literature to visually explore and understand the underlying key-insights using various information visualization techniques. We employ a natural language processing based pipeline to discover knowledge out of the biomedical dark data. The pipeline comprises of different lexical analysis techniques like Topic Modeling to extract inherent topics and major focus areas, Network Graphs to study the relationships between various entities like scientific documents and journals, researchers, and, keywords and terms, etc. With this analytical research, we aim to proffer a potential solution to overcome the problem of analyzing overwhelming amounts of information and diminish the limitation of human cognition and perception in handling and examining such large volumes of data.

READ FULL TEXT

page 4

page 7

page 8

page 11

page 13

research
09/04/2023

Learning a Patent-Informed Biomedical Knowledge Graph Reveals Technological Potential of Drug Repositioning Candidates

Drug repositioning-a promising strategy for discovering new therapeutic ...
research
08/17/2022

On the evolution of research in hypersonics: application of natural language processing and machine learning

Research and development in hypersonics have progressed significantly in...
research
05/22/2023

A Diachronic Analysis of the NLP Research Paradigm Shift: When, How, and Why?

Understanding the fundamental concepts and trends in a scientific field ...
research
02/03/2023

Analyzing the impact of climate change on critical infrastructure from the scientific literature: A weakly supervised NLP approach

Natural language processing (NLP) is a promising approach for analyzing ...
research
06/09/2022

Analyzing Folktales of Different Regions Using Topic Modeling and Clustering

This paper employs two major natural language processing techniques, top...
research
08/23/2020

Visual Exploration System for Analyzing Trends in Annual Recruitment Using Time-varying Graphs

Annual recruitment data of new graduates are manually analyzed by human ...
research
05/05/2020

A Pipeline for Integrated Theory and Data-Driven Modeling of Genomic and Clinical Data

High throughput genome sequencing technologies such as RNA-Seq and Micro...

Please sign up or login with your details

Forgot password? Click here to reset