Generative Adversarial Active Learning for Unsupervised Outlier Detection

09/28/2018
by   Yezheng Liu, et al.
18

Outlier detection is an important topic in machine learning and has been used in a wide range of applications. In this paper, we approach outlier detection as a binary-classification issue by sampling potential outliers from a uniform reference distribution. However, due to the sparsity of data in high-dimensional space, a limited number of potential outliers may fail to provide sufficient information to assist the classifier in describing a boundary that can separate outliers from normal data effectively. To address this, we propose a novel Single-Objective Generative Adversarial Active Learning (SO-GAAL) method for outlier detection, which can directly generate informative potential outliers based on the mini-max game between a generator and a discriminator. Moreover, to prevent the generator from falling into the mode collapsing problem, the stop node of training should be determined when SO-GAAL is able to provide sufficient information. But without any prior information, it is extremely difficult for SO-GAAL. Therefore, we expand the network structure of SO-GAAL from a single generator to multiple generators with different objectives (MO-GAAL), which can generate a reasonable reference distribution for the whole dataset. We empirically compare the proposed approach with several state-of-the-art outlier detection methods on both synthetic and real-world datasets. The results show that MO-GAAL outperforms its competitors in the majority of cases, especially for datasets with various cluster types or high irrelevant variable ratio.

READ FULL TEXT

page 5

page 6

page 9

page 12

research
04/21/2022

Fluctuation-based Outlier Detection

Outlier detection is an important topic in machine learning and has been...
research
06/28/2022

POEM: Out-of-Distribution Detection with Posterior Sampling

Out-of-distribution (OOD) detection is indispensable for machine learnin...
research
07/31/2019

Are Outlier Detection Methods Resilient to Sampling?

Outlier detection is a fundamental task in data mining and has many appl...
research
10/26/2018

Outlier Detection using Generative Models with Theoretical Performance Guarantees

This paper considers the problem of recovering signals from compressed m...
research
11/01/2022

Typical Yet Unlikely: Using Information Theoretic Approaches to Identify Outliers which Lie Close to the Mean

Normality, in the colloquial sense, has historically been considered an ...
research
09/30/2016

Outlier Detection from Network Data with Subnetwork Interpretation

Detecting a small number of outliers from a set of data observations is ...
research
03/07/2020

RCC-Dual-GAN: An Efficient Approach for Outlier Detection with Few Identified Anomalies

Outlier detection is an important task in data mining and many technolog...

Please sign up or login with your details

Forgot password? Click here to reset