Towards Sequence Utility Maximization under Utility Occupancy Measure

12/20/2022
by   Gengsen Huang, et al.
0

The discovery of utility-driven patterns is a useful and difficult research topic. It can extract significant and interesting information from specific and varied databases, increasing the value of the services provided. In practice, the measure of utility is often used to demonstrate the importance, profit, or risk of an object or a pattern. In the database, although utility is a flexible criterion for each pattern, it is a more absolute criterion due to the neglect of utility sharing. This leads to the derived patterns only exploring partial and local knowledge from a database. Utility occupancy is a recently proposed model that considers the problem of mining with high utility but low occupancy. However, existing studies are concentrated on itemsets that do not reveal the temporal relationship of object occurrences. Therefore, this paper towards sequence utility maximization. We first define utility occupancy on sequence data and raise the problem of High Utility-Occupancy Sequential Pattern Mining (HUOSPM). Three dimensions, including frequency, utility, and occupancy, are comprehensively evaluated in HUOSPM. An algorithm called Sequence Utility Maximization with Utility occupancy measure (SUMU) is proposed. Furthermore, two data structures for storing related information about a pattern, Utility-Occupancy-List-Chain (UOL-Chain) and Utility-Occupancy-Table (UO-Table) with six associated upper bounds, are designed to improve efficiency. Empirical experiments are carried out to evaluate the novel algorithm's efficiency and effectiveness. The influence of different upper bounds and pruning strategies is analyzed and discussed. The comprehensive results suggest that the work of our algorithm is intelligent and effective.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
12/28/2018

HUOPM: High Utility Occupancy Pattern Mining

Mining useful patterns from varied types of databases is an important re...
research
03/30/2021

TUSQ: Targeted High-Utility Sequence Querying

Significant efforts have been expended in the research and development o...
research
11/24/2021

Flexible Pattern Discovery and Analysis

Based on the analysis of the proportion of utility in the supporting tra...
research
08/26/2022

Itemset Utility Maximization with Correlation Measure

As an important data mining technology, high utility itemset mining (HUI...
research
11/26/2020

On-shelf Utility Mining of Sequence Data

Utility mining has emerged as an important and interesting topic owing t...
research
09/27/2022

Totally-ordered Sequential Rules for Utility Maximization

High utility sequential pattern mining (HUSPM) is a significant and valu...
research
08/18/2020

Discovering High Utility-Occupancy Patterns from Uncertain Data

It is widely known that there is a lot of useful information hidden in b...

Please sign up or login with your details

Forgot password? Click here to reset