Biomedical Document Clustering and Visualization based on the Concepts of Diseases

10/22/2018
by   Setu Shah, et al.
0

Document clustering is a text mining technique used to provide better document search and browsing in digital libraries or online corpora. A lot of research has been done on biomedical document clustering that is based on using existing ontology. But, associations and co-occurrences of the medical concepts are not well represented by using ontology. In this research, a vector representation of concepts of diseases and similarity measurement between concepts are proposed. They identify the closest concepts of diseases in the context of a corpus. Each document is represented by using the vector space model. A weight scheme is proposed to consider both local content and associations between concepts. A Self-Organizing Map is used as document clustering algorithm. The vector projection and visualization features of SOM enable visualization and analysis of the clusters distributions and relationships on the two dimensional space. The experimental results show that the proposed document clustering framework generates meaningful clusters and facilitate visualization of the clusters based on the concepts of diseases.

READ FULL TEXT

page 4

page 7

research
03/11/2020

ConceptScope: Organizing and Visualizing Knowledge in Documents based on Domain Ontology

Current text visualization techniques typically provide overviews of doc...
research
09/05/2018

Measures of Cluster Informativeness for Medical Evidence Aggregation and Dissemination

The largest collection of medical evidence in the world is PubMed. Howev...
research
11/03/2018

Learning Contextual Hierarchical Structure of Medical Concepts with Poincairé Embeddings to Clarify Phenotypes

Biomedical association studies are increasingly done using clinical conc...
research
01/08/2020

Techniques d'anonymisation tabulaire : concepts et mise en oeuvre

In this document, we present a state of the art of anonymization techniq...
research
04/02/2023

Enhancing Cluster Quality of Numerical Datasets with Domain Ontology

Ontology-based clustering has gained attention in recent years due to th...
research
03/02/2020

Cartolabe: A Web-Based Scalable Visualization of Large Document Collections

We describe CARTOLABE, a web-based multi-scale system for visualizing an...

Please sign up or login with your details

Forgot password? Click here to reset