节点文献

CAPE——数据流上的基于频繁模式的分类算法

CAPE—A Classification Algorithm Using Frequent Patterns over Data Streams

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王鹏吴晓晨王晨汪卫施伯乐

【Author】 WANG Peng,WU Xiao-Chen,WANG Chen,WANG Wei,and SHI Bai-Le (Department of Computing and Information Technology,Fudan University,Shanghai 200433)

【机构】 复旦大学计算机与信息技术系

【摘要】 近年来涌现出很多数据流的应用,比如网络日志、传感器网络等.数据流的数据量无限、数据分布变化等特性使得传统的挖掘算法不能很好地解决这些问题.针对上述问题提出了一种数据流上的基于频繁模式的分类算法——CAPE(classification using frequent pattern).CAPE通过数据流中的频繁模式进行分类,在压缩数据的同时保存了数据中的分类信息.实验证明,这种算法比其他算法有更高的准确性.并且CAPE可以很好地处理训练集包含大量缺失取值的应用.

【Abstract】 Classification is an important data mining task in the past decade.Meanwhile,many effective and efficient methods,e.g.decision tree and Bayes network,have been developed for classifying on large static database.However,these methods do not fit to processing over data stream.So a new algorithm—CAPE(classification using frequent patterns) is presented to deal with classification over data stream. Frequent patterns are imported into classification and used to record data distributing over stream mainly during a certain time slice.The experimental results show that the accuracy of classification using frequent patterns over stream is higher in most cases compared with the algorithm"weighted classifier ensembles"which is known to be the best classification algorithm over stream at present.

【关键词】 数据流分类决策树频繁模式
【Key words】 data streamclassificationdecision treefrequent pattern
【基金】 国家自然科学基金重点项目(69933010,60303008);国家“八六三”高技术研究发展计划基金项目(2002AA4Z3430,2002AA231041)
  • 【会议录名称】 第二十一届中国数据库学术会议论文集(研究报告篇)
  • 【会议名称】第二十一届中国数据库学术会议
  • 【会议时间】2004-10-14
  • 【会议地点】中国福建厦门
  • 【分类号】TP311.13
  • 【主办单位】中国计算机学会数据库专业委员会
节点文献中: 

本文链接的文献网络图示:

本文的引文网络