Explanations of Black-Box Models based on Directional Feature Interactions

04/16/2023
by   Aria Masoomi, et al.
0

As machine learning algorithms are deployed ubiquitously to a variety of domains, it is imperative to make these often black-box models transparent. Several recent works explain black-box models by capturing the most influential features for prediction per instance; such explanation methods are univariate, as they characterize importance per feature. We extend univariate explanation to a higher-order; this enhances explainability, as bivariate methods can capture feature interactions in black-box models, represented as a directed graph. Analyzing this graph enables us to discover groups of features that are equally important (i.e., interchangeable), while the notion of directionality allows us to identify the most influential features. We apply our bivariate method on Shapley value explanations, and experimentally demonstrate the ability of directional explanations to discover feature interactions. We show the superiority of our method against state-of-the-art on CIFAR10, IMDB, Census, Divorce, Drug, and gene data.

READ FULL TEXT

page 9

page 31

research
06/16/2020

High Dimensional Model Explanations: an Axiomatic Approach

Complex black-box machine learning models are regularly used in critical...
research
04/01/2019

VINE: Visualizing Statistical Interactions in Black Box Models

As machine learning becomes more pervasive, there is an urgent need for ...
research
03/03/2022

Label-Free Explainability for Unsupervised Models

Unsupervised black-box models are challenging to interpret. Indeed, most...
research
12/02/2019

Diagnostic Curves for Black Box Models

In safety-critical applications of machine learning, it is often necessa...
research
12/21/2018

Example and Feature importance-based Explanations for Black-box Machine Learning Models

As machine learning models become more accurate, they typically become m...
research
06/24/2022

Analyzing the Effects of Classifier Lipschitzness on Explainers

Machine learning methods are getting increasingly better at making predi...
research
06/01/2022

Composition of Relational Features with an Application to Explaining Black-Box Predictors

Relational machine learning programs like those developed in Inductive L...

Please sign up or login with your details

Forgot password? Click here to reset