Community-Based Hierarchical Positive-Unlabeled (PU) Model Fusion for Chronic Disease Prediction

09/06/2023
by   Yang Wu, et al.
0

Positive-Unlabeled (PU) Learning is a challenge presented by binary classification problems where there is an abundance of unlabeled data along with a small number of positive data instances, which can be used to address chronic disease screening problem. State-of-the-art PU learning methods have resulted in the development of various risk estimators, yet they neglect the differences among distinct populations. To address this issue, we present a novel Positive-Unlabeled Learning Tree (PUtree) algorithm. PUtree is designed to take into account communities such as different age or income brackets, in tasks of chronic disease prediction. We propose a novel approach for binary decision-making, which hierarchically builds community-based PU models and then aggregates their deliverables. Our method can explicate each PU model on the tree for the optimized non-leaf PU node splitting. Furthermore, a mask-recovery data augmentation strategy enables sufficient training of the model in individual communities. Additionally, the proposed approach includes an adversarial PU risk estimator to capture hierarchical PU-relationships, and a model fusion network that integrates data from each tree path, resulting in robust binary classification results. We demonstrate the superior performance of PUtree as well as its variants on two benchmarks and a new diabetes-prediction dataset.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
02/24/2020

Learning from Positive and Unlabeled Data with Arbitrary Positive Shift

Positive-unlabeled (PU) learning trains a binary classifier using only p...
research
03/14/2022

Improving State-of-the-Art in One-Class Classification by Leveraging Unlabeled Data

When dealing with binary classification of data with only one labeled cl...
research
05/02/2022

Positive-Unlabeled Learning with Adversarial Data Augmentation for Knowledge Graph Completion

Most real-world knowledge graphs (KG) are far from complete and comprehe...
research
11/25/2022

Positive unlabeled learning with tensor networks

Positive unlabeled learning is a binary classification problem with posi...
research
11/30/2022

Split-PU: Hardness-aware Training Strategy for Positive-Unlabeled Learning

Positive-Unlabeled (PU) learning aims to learn a model with rare positiv...
research
03/17/2023

Reliable Multimodality Eye Disease Screening via Mixture of Student's t Distributions

Multimodality eye disease screening is crucial in ophthalmology as it in...
research
05/19/2022

A Boosting Algorithm for Positive-Unlabeled Learning

Positive-unlabeled (PU) learning deals with binary classification proble...

Please sign up or login with your details

Forgot password? Click here to reset