Title
CCPM: A Scalable and Noise-Resistant Closed Contiguous Sequential Patterns Mining Algorithm.
Abstract
Mining closed contiguous sequential patterns has been addressed in the literature only recently, through the CCSpan algorithm. CCSpan mines a set of patterns that contains the same information than traditional sets of closed sequential patterns, while being more compact due to the contiguity. Although CCSpan outperforms closed sequential pattern mining algorithms in the general case, it does not scale well on large datasets with long sequences. Moreover, in the context of noisy datasets, the contiguity constraint prevents from mining a relevant result set. Inspired by BIDE, that has proven to be one of the most efficient closed sequential pattern mining algorithm, we propose CCPM that mines closed contiguous sequential patterns, while being scalable. Furthermore, CCPM introduces usable wildcards that address the problem of mining noisy data. Experiments show that CCPM greatly outperforms CCSpan, especially on large datasets with long sequences. In addition, they show that the wildcards allows to efficiently tackle the problem of noisy data.
Year
Venue
Field
2017
MLDM
USable,Data mining,Contiguity,Pattern recognition,Result set,Wildcard character,Computer science,Contiguity (probability theory),Artificial intelligence,Data mining algorithm,Sequential Pattern Mining,Scalability
DocType
Citations 
PageRank 
Conference
0
0.34
References 
Authors
18
3
Name
Order
Citations
PageRank
Yacine Abboud100.34
Anne Boyer210618.08
Armelle Brun313821.49