节点文献
prefix-hash-tree的插入、查找和重构算法
Insertion,Finding and Reconstruction Algorithms of Prefix-Hash-Tree
【机构】 华中科技大学计算机学院数据库与多媒体技术研究所; 中国电力财务有限公司华中分公司;
【摘要】 <正>在现有的众多文本分类方法中,关联分类以其较高的准确率和较快的训练时间而成为一种重要的自动文本分类方法。关联规则挖掘算法最初用于挖掘大型事务数据库中项与项间的有趣关系,其中最为著名的算法是Apriori算法和FP-tree方法,后来Bing Liu等在文[3]中通过改造Apriori算法,
【Abstract】 Among many existed text categorization approaches,association rule based document classification has aroused great attention as to its high accuracy and fast training time.In this paper,a special data structure called prefix-hash-tree and its relevant insertion,finding and reconstruction algorithms are designed.Experiment shows that unstructured Chinese text can be efficiently transformed into structured transaction data using this data structure.
【Key words】 Chinese text classification;
Prefix-hash-tree;
Transaction data;
- 【会议录名称】 第二十一届中国数据库学术会议论文集(技术报告篇)
- 【会议名称】第二十一届中国数据库学术会议
- 【会议时间】2004-10-14
- 【会议地点】中国福建厦门
- 【分类号】TP301.6
- 【主办单位】中国计算机学会数据库专业委员会