Discovering High Utility Episodes in Sequences

12/25/2019
by   Wensheng Gan, et al.
0

Sequence data, e.g., complex event sequence, is more commonly seen than other types of data (e.g., transaction data) in real-world applications. For the mining task from sequence data, several problems have been formulated, such as sequential pattern mining, episode mining, and sequential rule mining. As one of the fundamental problems, episode mining has often been studied. The common wisdom is that discovering frequent episodes is not useful enough. In this paper, we propose an efficient utility mining approach namely UMEpi: Utility Mining of high-utility Episodes from complex event sequence. We propose the concept of remaining utility of episode, and achieve a tighter upper bound, namely episode-weighted utilization (EWU), which will provide better pruning. Thus, the optimized EWU-based pruning strategies can achieve better improvements in mining efficiency. The search space of UMEpi w.r.t. a prefix-based lexicographic sequence tree is spanned and determined recursively for mining high-utility episodes, by prefix-spanning in a depth-first way. Finally, extensive experiments on four real-life datasets demonstrate that UMEpi can discover the complete high-utility episodes from complex event sequence, while the state-of-the-art algorithms fail to return the correct results. Furthermore, the improved variants of UMEpi significantly outperform the baseline in terms of execution time, memory consumption, and scalability.

READ FULL TEXT
research
11/29/2021

US-Rule: Discovering Utility-driven Sequential Rules

Utility-driven mining is an important task in data science and has many ...
research
10/30/2021

Utility-driven Mining of Contiguous Sequences

Recently, contiguous sequential pattern mining (CSPM) gained interest as...
research
06/28/2021

THUE: Discovering Top-K High Utility Episodes

Episode discovery from an event is a popular framework for data mining t...
research
03/30/2021

Explainable Fuzzy Utility Mining on Sequences

Fuzzy systems have good modeling capabilities in several data science sc...
research
02/25/2019

Utility Mining Across Multi-Dimensional Sequences

Knowledge extraction from database is the fundamental task in database a...
research
12/21/2022

A Projected Upper Bound for Mining High Utility Patterns from Interval-Based Event Sequences

High utility pattern mining is an interesting yet challenging problem. T...
research
08/26/2022

Temporal Fuzzy Utility Maximization with Remaining Measure

High utility itemset mining approaches discover hidden patterns from lar...

Please sign up or login with your details

Forgot password? Click here to reset