Breast and Colon Cancer Classification from Gene Expression Profiles Using Data Mining Techniques

04/06/2020
by   Mohamed Loey , et al.
0

Early detection of cancer increases the probability of recovery. This paper presents an intelligent decision support system (IDSS) for the early diagnosis of cancer based on gene expression profiles collected using DNA microarrays. Such datasets pose a challenge because of the small number of samples (no more than a few hundred) relative to the large number of genes (in the order of thousands). Therefore, a method of reducing the number of features (genes) that are not relevant to the disease of interest is necessary to avoid overfitting. The proposed methodology uses the information gain (IG) to select the most important features from the input patterns. Then, the selected features (genes) are reduced by applying the grey wolf optimization (GWO) algorithm. Finally, the methodology employs a support vector machine (SVM) classifier for cancer type classification. The proposed methodology was applied to two datasets (Breast and Colon) and was evaluated based on its classification accuracy, which is the most important performance measure in disease diagnosis. The experimental results indicate that the proposed methodology is able to enhance the stability of the classification accuracy as well as the feature selection.

READ FULL TEXT

page 3

page 5

page 8

page 10

page 11

page 13

page 14

page 15

research
02/24/2022

An Efficient Binary Harris Hawks Optimization based on Quantum SVM for Cancer Classification Tasks

Cancer classification based on gene expression increases early diagnosis...
research
05/27/2022

Gene selection from microarray expression data: A Multi-objective PSO with adaptive K-nearest neighborhood

Cancer detection is one of the key research topics in the medical field....
research
07/12/2012

Biogeography-Based Informative Gene Selection and Cancer Classification Using SVM and Random Forests

Microarray cancer gene expression data comprise of very high dimensions....
research
08/08/2020

Extended Particle Swarm Optimization (EPSO) for Feature Selection of High Dimensional Biomedical Data

This paper proposes a novel Extended Particle Swarm Optimization model (...
research
03/06/2022

A SVM Model for Candidate Y-chromosome Gene Discovery in Prostate Cancer

Prostate cancer is widely known to be one of the most common cancers amo...
research
08/26/2021

SVM Classifier on Chip for Melanoma Detection

Support Vector Machine (SVM) is a common classifier used for efficient c...
research
10/31/2021

Predicting Cancer Using Supervised Machine Learning: Mesothelioma

Background: Pleural Mesothelioma (PM) is an unusual, belligerent tumor t...

Please sign up or login with your details

Forgot password? Click here to reset