2003•Journal of Computer Research and DevelopmentRequires access

Fast Mining of Global Frequent Itemsets

Zhi Jie Sun

Open publisher page 9 citations

Abstract

Fast mining of global frequent itemsets is an important data mining problem in a distributed database environment Conventional mining algorithms employ the same framework as Apriori for global frequent itemsets However, candidate set generation is still costly, and the algorithms need repeatedly scan the database, especially when there exist prolific patterns and/or long patterns And communication overhead is costly by transmitting local frequent itemsets for global frequent itemsets In this paper, an algorithm FMAGF(fast mining algorithm of global frequent itmesets) in distributed database is proposed The idea of FMAGF is to only transmit conditional frequent pattern trees or conditional pattern bases but not to transmit a lot of local frequent itemsets; therefore, the algorithm uses far less communication overhead and improves efficiency of mining global frequent itemsets Theory analysis and experimental results show the feasibility and effectiveness of the algorithm

About this research paper

What this paper is about

Fast mining of global frequent itemsets is an important data mining problem in a distributed database environment Conventional mining algorithms employ the same framework as Apriori for global frequent itemsets However, candidate set generation is still costly, and the algorithms need repeatedly scan the database, especially when there exist prolific patterns and/or long patterns And communication overhead is costly by transmitting local frequent itemsets for global frequent itemsets In this paper, an algorithm FMAGF(fast mining algorithm of global frequent itmesets) in distributed database is proposed The idea of FMAGF is to only transmit conditional frequent pattern trees or conditional pattern bases but not to transmit a lot of local frequent itemsets; therefore, the algorithm uses far less communication overhead and improves efficiency of mining global frequent itemsets Theory analysis and experimental results show the feasibility and effectiveness of the algorithm

Why it matters

OpenAlex reports 9 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

Fast mining of global frequent itemsets is an important data mining problem in a distributed database environment Conventional mining algorithms employ the same framework as Apriori for global frequent itemsets However, candidate set generation is still costly, and the algorithms need repeatedly scan the database, especially when there exist prolific patterns and/or long patterns And communication overhead is costly by transmitting local frequent itemsets for global frequent itemsets In this paper, an algorithm FMAGF(fast mining algorithm of global frequent itmesets) in distributed database is proposed The idea of FMAGF is to only transmit conditional frequent pattern trees or conditional pattern bases but not to transmit a lot of local frequent itemsets; therefore, the algorithm uses far less communication overhead and improves efficiency of mining global frequent itemsets Theory analysis and experimental results show the feasibility and effectiveness of the algorithm

Key concepts: Computer science, Data mining, Overhead (engineering), Apriori algorithm, Association rule learning, A priori and a posteriori, GSP Algorithm, Set (abstract data type)

Related papers

Back to paper searchBrowse research topicsOriginal source
Fast Mining of Global Frequent Itemsets — Research Paper | ScholarLens