Ensuring Dataset Quality for Machine Learning Certification

11/03/2020
by   Sylvaine Picard, et al.
0

In this paper, we address the problem of dataset quality in the context of Machine Learning (ML)-based critical systems. We briefly analyse the applicability of some existing standards dealing with data and show that the specificities of the ML context are neither properly captured nor taken into ac-count. As a first answer to this concerning situation, we propose a dataset specification and verification process, and apply it on a signal recognition system from the railway domain. In addi-tion, we also give a list of recommendations for the collection and management of datasets. This work is one step towards the dataset engineering process that will be required for ML to be used on safety critical systems.

READ FULL TEXT
research
09/30/2022

Empowering the trustworthiness of ML-based critical systems through engineering activities

This paper reviews the entire engineering process of trustworthy Machine...
research
09/28/2022

Toward Certification of Machine-Learning Systems for Low Criticality Airborne Applications

The exceptional progress in the field of machine learning (ML) in recent...
research
08/05/2018

Using Machine Learning Safely in Automotive Software: An Assessment and Adaption of Software Process Requirements in ISO 26262

The use of machine learning (ML) is on the rise in many sectors of softw...
research
03/02/2020

Towards Probability-based Safety Verification of Systems with Components from Machine Learning

Machine learning (ML) has recently created many new success stories. Hen...
research
05/27/2020

SafeML: Safety Monitoring of Machine Learning Classifiers through Statistical Difference Measure

Ensuring safety and explainability of machine learning (ML) is a topic o...
research
11/08/2018

Satyam: Democratizing Groundtruth for Machine Vision

The democratization of machine learning (ML) has led to ML-based machine...
research
12/15/2021

Fix your Models by Fixing your Datasets

The quality of underlying training data is very crucial for building per...

Please sign up or login with your details

Forgot password? Click here to reset