节点文献

基于网页结构挖掘的信息提取

Extracting Information by Mining Structures of Web Pages

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 李媛耿桦张甍潘金贵

【Author】 LI Yuan GENG Hua ZHANG Meng PAN Jin Gui(State Key Laboratory for Novel Software Technology of Nanjing University,Multimedia Technology Institute of Nanjing University,Nanjing 210093)

【机构】 南京大学计算机软件新技术国家重点实验室南京大学计算机软件新技术国家重点实验室 南京210093南京210093

【摘要】 本文提出了两种细粒度的、基于网页结构挖掘的信息提取方法,比较了它们的优缺点,并给出了相应具体实现的性能测试和结果分析。

【Abstract】 To simplify the task of obtaining information from the vast number of information sources that are available on the WWW,we have developed two different methods to extract information of fine grain.This paper firstly de- scribes the principles of the two methods,which work by mining structures of Web pages,and then compares the ad- vantages and disadvantages of them.Finally,we test the performance of the two methods and analyze the experiment results.

  • 【文献出处】 计算机科学 ,Computer Science , 编辑部邮箱 ,2006年03期
  • 【分类号】TP311.13
  • 【被引频次】12
  • 【下载频次】256
节点文献中: 

本文链接的文献网络图示:

本文的引文网络