Rate of Change Analysis for Interestingness Measures

12/14/2017
by   Nandan Sudarsanam, et al.
0

The use of Association Rule Mining techniques in diverse contexts and domains has resulted in the creation of numerous interestingness measures. This, in turn, has motivated researchers to come up with various classification schemes for these measures. One popular approach to classify the objective measures is to assess the set of mathematical properties they satisfy in order to help practitioners select the right measure for a given problem. In this research, we discuss the insufficiency of the existing properties in literature to capture certain behaviors of interestingness measures. This motivates us to present a novel approach to analyze and classify measures. We refer to this as a rate of change analysis (RCA). In this analysis a measure is described by how it varies if there is a unit change in the frequency count (f_11,f_10,f_01,f_00), for different pre-existing states of the frequency counts. More formally, we look at the first partial derivative of the measure with respect to the various frequency count variables. We then use this analysis to define two new properties, Unit-Null Asymptotic Invariance (UNAI) and Unit-Null Zero Rate (UNZR). UNAI looks at the asymptotic effect of adding frequency patterns, while UNZR looks at the initial effect of adding frequency patterns when they do not pre-exist in the dataset. We present a comprehensive analysis of 50 interestingness measures and classify them in accordance with the two properties. We also present empirical studies, involving both synthetic and real-world datasets, which are used to cluster various measures according to the rule ranking patterns of the measures. The study concludes with the observation that classification of measures using the empirical clusters share significant similarities to the classification of measures done through the properties presented in this research.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/22/2022

Good Classification Measures and How to Find Them

Several performance measures can be used for evaluating classification r...
research
04/10/2015

Performance measures for classification systems with rejection

Classifiers with rejection are essential in real-world applications wher...
research
08/16/2013

Standardizing Interestingness Measures for Association Rules

Interestingness measures provide information that can be used to prune o...
research
04/24/2017

Visual-Based Analysis of Classification Measures with Applications to Imbalanced Data

With a plethora of available classification performance measures, choosi...
research
09/21/2016

On Data-Independent Properties for Density-Based Dissimilarity Measures in Hybrid Clustering

Hybrid clustering combines partitional and hierarchical clustering for c...
research
06/09/2023

Null/No Information Rate (NIR): a statistical test to assess if a classification accuracy is significant for a given problem

In many research contexts, especially in the biomedical field, after stu...
research
09/06/2018

Evaluation Measures for Quantification: An Axiomatic Approach

Quantification is the task of estimating, given a set σ of unlabelled it...

Please sign up or login with your details

Forgot password? Click here to reset