节点文献

基于SPAM-FPT的WebLog访问序列模式挖掘

WebLog Access Sequential Pattern Mining Based on SPAM-FTP

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 朱莉应吉康卜忠飞

【Author】 ZHU Li YING Ji-kang Computing Center of Information Institute,East China Normal University,Shanghai 200062 BU Zhong-fei Education Department of Yangzhou,Yangzhou 225002

【机构】 华东师范大学信息学院计算中心扬州市教育局 上海 200062上海 200062扬州 225002

【摘要】 WebLog访问序列模式挖掘将数据挖掘中的序列模式技术应用于Web服务器上的日志文件,以此来改善Web的信息服务,而在对海量的数据挖掘时,系统资源开销很大。该文结合SPAM、PrefixSpan的思想,提出一个新的算法——SPAM-FPT,该算法通过建立First_Positon_Table,避免了SPAM中的"与操作"、"连接操作"以及PrefixSpan中大量的"投影数据库"的建立,可以快捷地挖掘数据库中所有"频繁子序列"。

【Abstract】 WebLog mining is application of sequential pattern mining of data mining technology on Web server log files.Sequential patterns mined from Web logs are used to improve the quality of information service on Web.The main challenge of mining access sequential pattern form WebLog is the high processing cost due to the large amount of data.By combining SPAM and PrefixSpan,this paper proposes a new arithmetic SPAM-FPT.By constructing first_positon_table,SPAM-fPT avoids"joining"or"ANDing"in SPAM and generating a large number of projected database in PrefixSpan,and gets all the frequent sequential patterns form dtatabase.

  • 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2007年17期
  • 【分类号】TP393.09;TP311.13
  • 【被引频次】3
  • 【下载频次】122
节点文献中: 

本文链接的文献网络图示:

本文的引文网络