Itemset Utility Maximization with Correlation Measure

08/26/2022
by   Jiahui Chen, et al.
0

As an important data mining technology, high utility itemset mining (HUIM) is used to find out interesting but hidden information (e.g., profit and risk). HUIM has been widely applied in many application scenarios, such as market analysis, medical detection, and web click stream analysis. However, most previous HUIM approaches often ignore the relationship between items in an itemset. Therefore, many irrelevant combinations (e.g., {gold, apple} and {notebook, book}) are discovered in HUIM. To address this limitation, many algorithms have been proposed to mine correlated high utility itemsets (CoHUIs). In this paper, we propose a novel algorithm called the Itemset Utility Maximization with Correlation Measure (CoIUM), which considers both a strong correlation and the profitable values of the items. Besides, the novel algorithm adopts a database projection mechanism to reduce the cost of database scanning. Moreover, two upper bounds and four pruning strategies are utilized to effectively prune the search space. And a concise array-based structure named utility-bin is used to calculate and store the adopted upper bounds in linear time and space. Finally, extensive experimental results on dense and sparse datasets demonstrate that CoIUM significantly outperforms the state-of-the-art algorithms in terms of runtime and memory consumption.

READ FULL TEXT
research
12/29/2022

HUSP-SP: Faster Utility Mining on Sequence Data

High-utility sequential pattern mining (HUSPM) has emerged as an importa...
research
06/28/2021

TOPIC: Top-k High-Utility Itemset Discovering

Utility-driven itemset mining is widely applied in many real-world scena...
research
04/16/2019

ProUM: Projection-based Utility Mining on Sequence Data

In recent decade, utility mining has attracted a great attention, but mo...
research
12/20/2022

Towards Sequence Utility Maximization under Utility Occupancy Measure

The discovery of utility-driven patterns is a useful and difficult resea...
research
08/26/2022

Temporal Fuzzy Utility Maximization with Remaining Measure

High utility itemset mining approaches discover hidden patterns from lar...
research
12/18/2018

High-utility itemset mining for subadditive monotone utility functions

High-utility Itemset Mining (HUIM) finds itemsets from a transaction dat...
research
03/30/2021

TUSQ: Targeted High-Utility Sequence Querying

Significant efforts have been expended in the research and development o...

Please sign up or login with your details

Forgot password? Click here to reset