Feature Selection Based on Unique Relevant Information for Health Data

12/02/2018
by   Shiyu Liu, et al.
0

Feature selection, which searches for the most representative features in observed data, is critical for health data analysis. Unlike feature extraction, such as PCA and autoencoder based methods, feature selection preserves interpretability, meaning that the selected features provide direct information about certain health conditions (i.e., the label). Thus, feature selection allows domain experts, such as clinicians, to understand the predictions made by machine learning based systems, as well as improve their own diagnostic skills. Mutual information is often used as a basis for feature selection since it measures dependencies between features and labels. In this paper, we introduce a novel mutual information based feature selection (MIBFS) method called SURI, which boosts features with high unique relevant information. We compare SURI to existing MIBFS methods using 3 different classifiers on 6 publicly available healthcare data sets. The results indicate that, in addition to preserving interpretability, SURI selects more relevant feature subsets which lead to higher classification performance. More importantly, we explore the dynamics of mutual information on a public low-dimensional health data set via exhaustive search. The results suggest the important role of unique relevant information in feature selection and verify the principles behind SURI.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/13/2020

Active Feature Selection for the Mutual Information Criterion

We study active feature selection, a novel feature selection setting in ...
research
12/09/2022

Improving Mutual Information based Feature Selection by Boosting Unique Relevance

Mutual Information (MI) based feature selection makes use of MI to evalu...
research
10/26/2021

Single Morphing Attack Detection using Feature Selection and Visualisation based on Mutual Information

Face morphing attack detection is a challenging task. Automatic classifi...
research
04/06/2023

SLM: End-to-end Feature Selection via Sparse Learnable Masks

Feature selection has been widely used to alleviate compute requirements...
research
08/06/2022

Efficient Novelty Detection Methods for Early Warning of Potential Fatal Diseases

Fatal diseases, as Critical Health Episodes (CHEs), represent real dange...
research
02/21/2023

Feature selection algorithm based on incremental mutual information and cockroach swarm optimization

Feature selection is an effective preprocessing technique to reduce data...
research
01/10/2021

Curvature-based Feature Selection with Application in Classifying Electronic Health Records

Electronic Health Records (EHRs) are widely applied in healthcare facili...

Please sign up or login with your details

Forgot password? Click here to reset