节点文献

挖掘关联规则算法的优化处理

Algorithm Optimization of Mining Association Rules

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 陆丽娜xjtu.edu.cn陈亚萍杨麦顺魏恒义

【Author】 Lu Lina; Chen Yaping; Yang Maishun; Wei Hengyi(Department of Computer Science and Technology,Xi’an Jiaotong University,Xi,an 710049)

【机构】 西安交通大学计算机科学与技术系!西安710049E-mail:lnlu

【摘要】 在挖掘关联规则的执行过程中,早期循环生成最大项目集的过程是很重要的。文中提出基于哈希表的算法,对生成侯选项目集的过程进行了优化,尤其是对生成二维侯选项目集更是有效。由于在早期循环中,生成侯选项目集的势较小,使得能更有效地修剪数据库,从而减小了后期循环的计算代价,同时也减小了I/O请求。

【Abstract】 To find all the large itemsets from candidate sets in eary iterations is usually the domaining factor foroverall data-mining performance. In the paper,we option the algorithm Apriori for the candidate set generation. It is ahash-based algorithm and is especially effective for the generation of candidate set for large 2-itemsets. Furthermore thegeneration of smaller candidate sets enables us to effectively trim the transaction database size at a much earier stageof the iterations,thereby reducing the computational cost for later iterations. The advantage of proposed algorithm alsoprovides us the opportunity of reducing the amount of disk I/O required.

【关键词】 数据挖掘关联规则哈希
【Key words】 Data MiningAssociation RulesHashing
  • 【文献出处】 计算机工程与应用 ,COMPUTER ENGINEERING AND APPLICATIONS , 编辑部邮箱 ,2000年08期
  • 【分类号】TP311
  • 【被引频次】42
  • 【下载频次】209
节点文献中: 

本文链接的文献网络图示:

本文的引文网络