FAIRification of MLC data

11/23/2022
by   Ana Kostovska, et al.
0

The multi-label classification (MLC) task has increasingly been receiving interest from the machine learning (ML) community, as evidenced by the growing number of papers and methods that appear in the literature. Hence, ensuring proper, correct, robust, and trustworthy benchmarking is of utmost importance for the further development of the field. We believe that this can be achieved by adhering to the recently emerged data management standards, such as the FAIR (Findable, Accessible, Interoperable, and Reusable) and TRUST (Transparency, Responsibility, User focus, Sustainability, and Technology) principles. To FAIRify the MLC datasets, we introduce an ontology-based online catalogue of MLC datasets that follow these principles. The catalogue extensively describes many MLC datasets with comprehensible meta-features, MLC-specific semantic descriptions, and different data provenance information. The MLC data catalogue is extensively described in our recent publication in Nature Scientific Reports, Kostovska Bogatinovski et al., and available at: http://semantichub.ijs.si/MLCdatasets. In addition, we provide an ontology-based system for easy access and querying of performance/benchmark data obtained from a comprehensive MLC benchmark study. The system is available at: http://semantichub.ijs.si/MLCbenchmark.

READ FULL TEXT
research
01/10/2022

FAIR high level data for Cherenkov astronomy

We highlight here several solutions developed to make high-level Cherenk...
research
07/26/2019

Making Neural Networks FAIR

Research on neural networks has gained significant momentum over the pas...
research
02/14/2021

Comprehensive Comparative Study of Multi-Label Classification Methods

Multi-label classification (MLC) has recently received increasing intere...
research
09/05/2018

Theoretical analysis and propositions for "ontology citation"

Ontology citation, the practice of referring the ontology in a similar f...
research
09/06/2022

Use and Misuse of Machine Learning in Anthropology

Machine learning (ML), being now widely accessible to the research commu...
research
07/14/2021

The I-ADOPT Interoperability Framework for FAIRer data descriptions of biodiversity

Biodiversity, the variation within and between species and ecosystems, i...
research
11/16/2021

GAP Enhancing Semantic Interoperability of Genomic Datasets and Provenance Through Nanopublications

While the publication of datasets in scientific repositories has become ...

Please sign up or login with your details

Forgot password? Click here to reset