节点文献

一种基于FP-Tree的频繁模式挖掘自适应算法

A Self-Adaptive Algorithm Based on FP-Tree for Frequent Pattern Mining

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 张锦马海兵胡运发

【Author】 ZHANG Jin;MA Hai-Bing;HU Yun-Fa Department of Computing and Information Technology, Fudan University,Shanghai 200433

【机构】 复旦大学计算机与信息技术系

【摘要】 不同数据集中数据的不同分布特征,对于频繁模式挖掘算法往往有着较大影响。将不同的现有算法结合起来,根据数据集的不同特性采用不同的挖掘策略,有可能构造出鲁棒性强的新算法。本文首先提出了一种基于FP-tree的简单深度优先搜索算法NDFS,并简单分析了其在不同数据集上的特性。在分析的基础上,本文进一步将NDFS和经典的FP-growth算法进行结合,提出了一种在挖掘过程中根据局部空间特征动态采用不同策略的自适应算法SAFP。实验证明,SAFP算法在不同数据集上均能达到或优于原有最优算法的性能,具有较好的鲁棒性。

【Abstract】 The distinct characters of different datasets greatly influence the efficiency of specific methods in frequent pattern mining. It is possible to build a robust algorithm by methodically combining different algorithms that should be properly applied according to the characters of data distribution of current dataset. This paper firstly proposes the Naive Depth First Search algorithm (NDFS) that is based on FP-tree, and then briefly analyzes its performance on different datasets. Finally, a new self-adaptive algorithm (SAFP) is proposed, which combines the NDFS with the FP-growth by a dynamic mining strategy on conditional FP-trees. Experiments demonstrate that the SAFP is more robust and efficient than both the NDFS and the FP-growth on various datasets.

【基金】 国家863高技术研究发展计划资助项目(No.2002AA1Z6707)
  • 【文献出处】 模式识别与人工智能 ,Pattern Recognition and Artificial Intelligence , 编辑部邮箱 ,2005年06期
  • 【分类号】TP18
  • 【被引频次】5
  • 【下载频次】63
节点文献中: 

本文链接的文献网络图示:

本文的引文网络