Parallel Mining of Maximal Frequent Itemsets from Databases

Authors:
Soon M. Chung;Congnan Luo
Affiliations:
-;-
Venue:
ICTAI '03 Proceedings of the 15th IEEE International Conference on Tools with Artificial Intelligence
Year:
2003

Citing 0
Cited 3

Parallel Leap: Large-Scale Maximal Pattern Mining in a Distributed Environment

ICPADS '06 Proceedings of the 12th International Conference on Parallel and Distributed Systems - Volume 1
Parallel Bifold: Large-scale parallel pattern mining with constraints

Distributed and Parallel Databases
A parallel algorithm for mining multiple partial periodic patterns

Information Sciences: an International Journal

Quantified Score

Hi-index	0.00

Visualization

Abstract

In this paper, we propose a parallel algorithm for mining maximal frequent itemsets from databases. A frequent itemset is maximal if none of its supersets is frequent. The new parallel algorithm is named Parallel Max-Miner (PMM), and it is a parallel version of the sequential Max-Miner algorithm [3]. Most of existing mining algorithms discover the frequent k-itemsets on the kth pass over the databases, and then generate the candidate (k + 1)-itemsets for the next pass. Compared to those level-wise algorithms, PMM looks ahead at each pass and prunes more candidate itemsets by checking the frequences of their supersets. We implemented PMM on a cluster of workstations, and evaluated its performance for various cases. PMM demonstrated better performance than other sequential and parallel algorithms, and its performance is quite scalable, even when there are large maximal frequent itemsets (i.e., long patterns) in databases.