An Efficient Anomaly Detection Approach using Cube Sampling with Streaming Data

10/05/2021
by   Seemandhar Jain, et al.
4

Anomaly detection is critical in various fields, including intrusion detection, health monitoring, fault diagnosis, and sensor network event detection. The isolation forest (or iForest) approach is a well-known technique for detecting anomalies. It is, however, ineffective when dealing with dynamic streaming data, which is becoming increasingly prevalent in a wide variety of application areas these days. In this work, we extend our previous work by proposed an efficient iForest based approach for anomaly detection using cube sampling that is effective on streaming data. Cube sampling is used in the initial stage to choose nearly balanced samples, significantly reducing storage requirements while preserving efficiency. Following that, the streaming nature of data is addressed by a sliding window technique that generates consecutive chunks of data for systematic processing. The novelty of this paper is in applying Cube sampling in iForest and calculating inclusion probability. The proposed approach is equally successful at detecting anomalies as existing state-of-the-art approaches, requiring significantly less storage and time complexity. We undertake empirical evaluations of the proposed approach using standard datasets and demonstrate that it outperforms traditional approaches in terms of Area Under the ROC Curve (AUC-ROC) and can handle high-dimensional streaming data.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/31/2022

AVTPnet: Convolutional Autoencoder for AVTP anomaly detection in Automotive Ethernet Networks

Network Intrusion Detection Systems are well considered as efficient too...
research
09/05/2022

RX-ADS: Interpretable Anomaly Detection using Adversarial ML for Electric Vehicle CAN data

Recent year has brought considerable advancements in Electric Vehicles (...
research
04/27/2021

Extending Isolation Forest for Anomaly Detection in Big Data via K-Means

Industrial Information Technology (IT) infrastructures are often vulnera...
research
05/03/2022

TracInAD: Measuring Influence for Anomaly Detection

As with many other tasks, neural networks prove very effective for anoma...
research
12/19/2018

Correlated Anomaly Detection from Large Streaming Data

Correlated anomaly detection (CAD) from streaming data is a type of grou...
research
11/13/2019

Real-Time Anomaly Detection for Advanced Manufacturing: Improving on Twitter's State of the Art

The detection of anomalies in real time is paramount to maintain perform...

Please sign up or login with your details

Forgot password? Click here to reset