Ethical behavior in humans and machines – Evaluating training data quality for beneficial machine learning

by   Thilo Hagendorff, et al.

Machine behavior that is based on learning algorithms can be significantly influenced by the exposure to data of different qualities. Up to now, those qualities are solely measured in technical terms, but not in ethical ones, despite the significant role of training and annotation data in supervised machine learning. This is the first study to fill this gap by describing new dimensions of data quality for supervised machine learning applications. Based on the rationale that different social and psychological backgrounds of individuals correlate in practice with different modes of human-computer-interaction, the paper describes from an ethical perspective how varying qualities of behavioral data that individuals leave behind while using digital technologies have socially relevant ramification for the development of machine learning applications. The specific objective of this study is to describe how training data can be selected according to ethical assessments of the behavior it originates from, establishing an innovative filter regime to transition from the big data rationale n = all to a more selective way of processing data for training sets in machine learning. The overarching aim of this research is to promote methods for achieving beneficial machine learning applications that could be widely useful for industry as well as academia.



There are no comments yet.


page 1

page 2

page 3

page 4


Can Machine Learning be Moral?

The ethics of Machine Learning has become an unavoidable topic in the AI...

When is it right and good for an intelligent autonomous vehicle to take over control (and hand it back)?

There is much debate in machine ethics about the most appropriate way to...

Consider ethical and social challenges in smart grid research

Artificial Intelligence and Machine Learning are increasingly seen as ke...

Forbidden knowledge in machine learning – Reflections on the limits of research and publication

Certain research strands can yield "forbidden knowledge". This term refe...

Human's Role in-the-Loop

Data integration has been recently challenged by the need to handle larg...

Big data ethics, machine ethics or information ethics? Navigating the maze of applied ethics in IT

Digitalization efforts are rapidly spreading across societies, challengi...

Uniqueness of Medical Data Mining: How the new technologies and data they generate are transforming medicine

The paper describes how the new technologies and data they generate are ...
This week in AI

Get the week's most popular data science and artificial intelligence research sent straight to your inbox every Saturday.