DOI | Trouver le DOI : https://doi.org/10.1109/BigData.2018.8622138 |
---|
Auteur | Rechercher : Buffett, Scott1 |
---|
Affiliation | - Conseil national de recherches du Canada. Technologies numériques
|
---|
Format | Texte, Article |
---|
Conférence | 2018 IEEE International Conference on Big Data (Big Data), December 10-13, 2018, Seattle, WA, USA |
---|
Sujet | high utility sequential pattern mining; sequential pattern mining; frequent pattern mining; candidate list maintenance |
---|
Résumé | High utility sequential pattern mining (HUSPM) lends the aspect of item value or importance to sequential pattern mining by identifying patterns that comprise a significant level of utility in a database. This paper addresses the challenge of establishing upper bounds on future candidate pattern utilities in an effort to reduce the search space required to identify the full set of patterns, and proposes a new approach where a list of possible candidate concatenation items is maintained. This list specifies the only items that ever need to be considered as possible candidates for concatenation with a sequential pattern being considered, or any future sequential pattern appearing as a descendant in the search tree. As a result of the elimination of items that are known to have no possibility of appearing in future high utility sequential patterns, an approach is presented that exploits this knowledge and computes a significantly tighter upper bound on the utilities of the such patterns. Tests on a variety of publicly available datasets show a dramatic reduction in the number of candidates considered, and the time taken to identify the full set of high utility sequential patterns is significantly reduced accordingly. |
---|
Date de publication | 2019-01-24 |
---|
Maison d’édition | IEEE |
---|
Dans | |
---|
Langue | anglais |
---|
Publications évaluées par des pairs | Oui |
---|
Exporter la notice | Exporter en format RIS |
---|
Signaler une correction | Signaler une correction (s'ouvre dans un nouvel onglet) |
---|
Identificateur de l’enregistrement | a0ed1a26-3195-4696-8254-6c2103e857e8 |
---|
Enregistrement créé | 2019-05-14 |
---|
Enregistrement modifié | 2022-02-21 |
---|