AIDA: Analytic Isolation and Distance-based Anomaly Detection Algorithm

12/05/2022
by   Luis Antonio Souto Arias, et al.
0

We combine the metrics of distance and isolation to develop the Analytic Isolation and Distance-based Anomaly (AIDA) detection algorithm. AIDA is the first distance-based method that does not rely on the concept of nearest-neighbours, making it a parameter-free model. Differently from the prevailing literature, in which the isolation metric is always computed via simulations, we show that AIDA admits an analytical expression for the outlier score, providing new insights into the isolation metric. Additionally, we present an anomaly explanation method based on AIDA, the Tempered Isolation-based eXplanation (TIX) algorithm, which finds the most relevant outlier features even in data sets with hundreds of dimensions. We test both algorithms on synthetic and empirical data: we show that AIDA is competitive when compared to other state-of-the-art methods, and it is superior in finding outliers hidden in multidimensional feature subspaces. Finally, we illustrate how the TIX algorithm is able to find outliers in multidimensional feature subspaces, and use these explanations to analyze common benchmarks used in anomaly detection.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/13/2021

Why Are You Weird? Infusing Interpretability in Isolation Forest for Anomaly Detection

Anomaly detection is concerned with identifying examples in a dataset th...
research
04/09/2020

A Mathematical Assessment of the Isolation Tree Method for Data Anomaly Detection in Big Data

We present the mathematical analysis of the Isolation Random Forest Meth...
research
10/26/2021

Revisiting randomized choices in isolation forests

Isolation forest or "iForest" is an intuitive and widely used algorithm ...
research
10/27/2019

Distance approximation using Isolation Forests

This work briefly explores the possibility of approximating spatial dist...
research
05/15/2023

A Sweep-plane Algorithm for Calculating the Isolation of Mountains

One established metric to classify the significance of a mountain peak i...
research
01/14/2019

CFOF: A Concentration Free Measure for Anomaly Detection

We present a novel notion of outlier, called the Concentration Free Outl...
research
06/01/2014

l_1-regularized Outlier Isolation and Regression

This paper proposed a new regression model called l_1-regularized outlie...

Please sign up or login with your details

Forgot password? Click here to reset