Statistical Analysis of Nearest Neighbor Methods for Anomaly Detection

07/08/2019
by   Xiaoyi Gu, et al.
0

Nearest-neighbor (NN) procedures are well studied and widely used in both supervised and unsupervised learning problems. In this paper we are concerned with investigating the performance of NN-based methods for anomaly detection. We first show through extensive simulations that NN methods compare favorably to some of the other state-of-the-art algorithms for anomaly detection based on a set of benchmark synthetic datasets. We further consider the performance of NN methods on real datasets, and relate it to the dimensionality of the problem. Next, we analyze the theoretical properties of NN-methods for anomaly detection by studying a more general quantity called distance-to-measure (DTM), originally developed in the literature on robust geometric and topological inference. We provide finite-sample uniform guarantees for the empirical DTM and use them to derive misclassification rates for anomalous observations under various settings. In our analysis we rely on Huber's contamination model and formulate mild geometric regularity assumptions on the underlying distribution of the data.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
04/27/2019

Guarantees on Nearest-Neighbor Condensation heuristics

The problem of nearest-neighbor (NN) condensation aims to reduce the siz...
research
05/01/2023

Unsupervised anomaly detection algorithms on real-world data: how many do we need?

In this study we evaluate 32 unsupervised anomaly detection algorithms o...
research
05/05/2020

Sub-Image Anomaly Detection with Deep Pyramid Correspondences

Nearest neighbor (kNN) methods utilizing deep pre-trained features exhib...
research
02/06/2015

Learning Efficient Anomaly Detectors from K-NN Graphs

We propose a non-parametric anomaly detection algorithm for high dimensi...
research
09/02/2019

An Adjusted Nearest Neighbor Algorithm Maximizing the F-Measure from Imbalanced Data

In this paper, we address the challenging problem of learning from imbal...
research
07/24/2022

e-G2C: A 0.14-to-8.31 μJ/Inference NN-based Processor with Continuous On-chip Adaptation for Anomaly Detection and ECG Conversion from EGM

This work presents the first silicon-validated dedicated EGM-to-ECG (G2C...
research
07/19/2018

Anomaly Detection for Water Treatment System based on Neural Network with Automatic Architecture Optimization

We continue to develop our neural network (NN) based forecasting approac...

Please sign up or login with your details

Forgot password? Click here to reset