Tips, guidelines and tools for managing multi-label datasets: the mldr.datasets R package and the Cometa data repository

02/10/2018
by   Francisco Charte, et al.
0

New proposals in the field of multi-label learning algorithms have been growing in number steadily over the last few years. The experimentation associated with each of them always goes through the same phases: selection of datasets, partitioning, training, analysis of results and, finally, comparison with existing methods. This last step is often hampered since it involves using exactly the same datasets, partitioned in the same way and using the same validation strategy. In this paper we present a set of tools whose objective is to facilitate the management of multi-label datasets, aiming to standardize the experimentation procedure. The two main tools are an R package, mldr.datasets, and a web repository with datasets, Cometa. Together, these tools will simplify the collection of datasets, their partitioning, documentation and export to multiple formats, among other functions. Some tips, recommendations and guidelines for a good experimental analysis of multi-label methods are also presented.

READ FULL TEXT

page 5

page 16

page 17

research
05/12/2020

Unsupervised Multi-label Dataset Generation from Web Data

This paper presents a system towards the generation of multi-label datas...
research
05/02/2019

Synthetic Oversampling of Multi-Label Data based on Local Label Distribution

Class-imbalance is an inherent characteristic of multi-label data which ...
research
12/25/2015

Inducing Generalized Multi-Label Rules with Learning Classifier Systems

In recent years, multi-label classification has attracted a significant ...
research
03/21/2019

Semantic Comparison of State-of-the-Art Deep Learning Methods for Image Multi-Label Classification

Image understanding relies heavily on accurate multi-label classificatio...
research
10/08/2022

A Survey on Extreme Multi-label Learning

Multi-label learning has attracted significant attention from both acade...
research
02/04/2022

Structured Prediction Problem Archive

Structured prediction problems are one of the fundamental tools in machi...

Please sign up or login with your details

Forgot password? Click here to reset