Demonstrating Rosa: the fairness solution for any Data Analytic pipeline

02/28/2020
by   Kate Wilkinson, et al.
0

Most datasets of interest to the analytics industry are impacted by various forms of human bias. The outcomes of Data Analytics [DA] or Machine Learning [ML] on such data are therefore prone to replicating the bias. As a result, a large number of biased decision-making systems based on DA/ML have recently attracted attention. In this paper we introduce Rosa, a free, web-based tool to easily de-bias datasets with respect to a chosen characteristic. Rosa is based on the principles of Fair Adversarial Networks, developed by illumr Ltd., and can therefore remove interactive, non-linear, and non-binary bias. Rosa is stand-alone pre-processing step / API, meaning it can be used easily with any DA/ML pipeline. We test the efficacy of Rosa in removing bias from data-driven decision making systems by performing standard DA tasks on five real-world datasets, selected for their relevance to current DA problems, and also their high potential for bias. We use simple ML models to model a characteristic of analytical interest, and compare the level of bias in the model output both with and without Rosa as a pre-processing step. We find that in all cases there is a substantial decrease in bias of the data-driven decision making systems when the data is pre-processed with Rosa.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
02/23/2020

Fair Adversarial Networks

The influence of human judgement is ubiquitous in datasets used across t...
research
08/01/2019

FairSight: Visual Analytics for Fairness in Decision Making

Data-driven decision making related to individuals has become increasing...
research
03/20/2023

Fairness-Aware Graph Filter Design

Graphs are mathematical tools that can be used to represent complex real...
research
05/21/2022

Automated machine learning: AI-driven decision making in business analytics

The realization that AI-driven decision-making is indispensable in today...
research
04/11/2017

Optimized Data Pre-Processing for Discrimination Prevention

Non-discrimination is a recognized objective in algorithmic decision mak...
research
05/22/2022

A Domain-adaptive Pre-training Approach for Language Bias Detection in News

Media bias is a multi-faceted construct influencing individual behavior ...
research
11/04/2022

Uncertainty-aware predictive modeling for fair data-driven decisions

Both industry and academia have made considerable progress in developing...

Please sign up or login with your details

Forgot password? Click here to reset