A Topological Approach for Semi-Supervised Learning

05/19/2022
by   Adrián Inés, et al.
30

Nowadays, Machine Learning and Deep Learning methods have become the state-of-the-art approach to solve data classification tasks. In order to use those methods, it is necessary to acquire and label a considerable amount of data; however, this is not straightforward in some fields, since data annotation is time consuming and might require expert knowledge. This challenge can be tackled by means of semi-supervised learning methods that take advantage of both labelled and unlabelled data. In this work, we present new semi-supervised learning methods based on techniques from Topological Data Analysis (TDA), a field that is gaining importance for analysing large amounts of data with high variety and dimensionality. In particular, we have created two semi-supervised learning methods following two different topological approaches. In the former, we have used a homological approach that consists in studying the persistence diagrams associated with the data using the Bottleneck and Wasserstein distances. In the latter, we have taken into account the connectivity of the data. In addition, we have carried out a thorough analysis of the developed methods using 3 synthetic datasets, 5 structured datasets, and 2 datasets of images. The results show that the semi-supervised methods developed in this work outperform both the results obtained with models trained with only manually labelled data, and those obtained with classical semi-supervised learning methods, reaching improvements of up to a 16

READ FULL TEXT

page 1

page 2

page 3

page 4

research
06/09/2020

An Overview of Deep Semi-Supervised Learning

Deep neural networks demonstrated their ability to provide remarkable pe...
research
05/16/2019

Dealing with Label Scarcity in Computational Pathology: A Use Case in Prostate Cancer Classification

Large amounts of unlabelled data are commonplace for many applications i...
research
03/24/2019

Exploiting Synthetically Generated Data with Semi-Supervised Learning for Small and Imbalanced Datasets

Data augmentation is rapidly gaining attention in machine learning. Synt...
research
11/26/2018

HELOC Applicant Risk Performance Evaluation by Topological Hierarchical Decomposition

Strong regulations in the financial industry mean that any decisions bas...
research
05/19/2022

Semi-Supervised Learning for Image Classification using Compact Networks in the BioMedical Context

The development of mobile and on the edge applications that embed deep c...
research
09/20/2017

Supervised Learning with Indefinite Topological Kernels

Topological Data Analysis (TDA) is a recent and growing branch of statis...
research
03/14/2022

Don't fear the unlabelled: safe deep semi-supervised learning via simple debiasing

Semi supervised learning (SSL) provides an effective means of leveraging...

Please sign up or login with your details

Forgot password? Click here to reset