Cluster ComputingEngineeringFrequent Pattern MiningData ScienceData MiningAssociation RuleMapreduce ImplementationsKnowledge DiscoveryFrequent Itemset MiningPattern DiscoveryPattern MiningData IntegrationMap-reduceMining MethodsData ManagementBig Data
One of the most important problems in data mining is frequent itemset mining. It requires very large computation and I/O traffic capacity. For that reason several parallel and distributed mining algorithms were developed. Recently the mapreduce distributed data processing paradigm is unavoidable and porting the current algorithms to mapreduce is in focus. In this paper a substantial frequent itemset mining algorithms and their mapreduce implementations are introduced and investi-gated. An algorithm improvement is also proposed and analyzed.
17
Fast algorithms for mining association rules
Rakesh Agrawal, Ramakrishnan Srikant · 1998 · 10.7K citations
Mining frequent patterns without candidate generation
Jiawei Han, Jian Pei, Yiwen Yin · 2000 · 3.2K citations · Full text