CCPM: A Scalable and Noise-Resistant Closed Contiguous Sequential Patterns Mining Algorithm - Archive ouverte HAL Access content directly
Conference Papers Year : 2017

CCPM: A Scalable and Noise-Resistant Closed Contiguous Sequential Patterns Mining Algorithm

(1) , (1) , (1)
1

Abstract

Mining closed contiguous sequential patterns has been addressed in the literature only recently, through the CCSpan algorithm. CCSpan mines a set of patterns that contains the same information than traditional sets of closed sequential patterns, while being more compact due to the contiguity. Although CCSpan outperforms closed sequential pattern mining algorithms in the general case, it does not scale well on large datasets with long sequences. Moreover, in the context of noisy datasets, the contiguity constraint prevents from mining a relevant result set. Inspired by BIDE, that has proven to be one of the most efficient closed sequential pattern mining algorithm, we propose CCPM that mines closed contiguous sequential patterns, while being scalable. Furthermore , CCPM introduces usable wildcards that address the problem of mining noisy data. Experiments show that CCPM greatly outperforms CCSpan, especially on large datasets with long sequences. In addition, they show that the wildcards allows to efficiently tackle the problem of noisy data.
Fichier principal
Vignette du fichier
CCPM A Scalable and Noise-Resistant Closed Contiguous Sequential Pattern....pdf (331.09 Ko) Télécharger le fichier
Origin : Files produced by the author(s)
Loading...

Dates and versions

hal-01569008 , version 1 (26-07-2017)

Identifiers

Cite

Yacine Abboud, Anne Boyer, Armelle Brun. CCPM: A Scalable and Noise-Resistant Closed Contiguous Sequential Patterns Mining Algorithm. 13th International Conference on Machine Learning and Data Mining MLDM 2017, Jul 2017, New York, United States. pp.15, ⟨10.1016/j.knosys.2015.06.014⟩. ⟨hal-01569008⟩
216 View
411 Download

Altmetric

Share

Gmail Facebook Twitter LinkedIn More