节点文献

一种基于XML的Web信息抽取方法

Study on Information Extraction Technology Based on Web Described with XML

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【摘要】 利用标准的XML技术来解决信息抽取问题,提出一个基于XML技术的Web信息抽取平台。通过归纳学习算法,寻找和识别出感兴趣的数据。利用XSLT和Xpath技术在数据定位和转换方面的优势,解决信息抽取中的关键问题:编写抽取规则。并对抽取规则进行优化,使其更加简单、健壮和通用。

【Abstract】 The standard XML technology be used to solve the information extraction problem in this article, proposed Web information extracted platform based on XML technical. And through the induction study algorithm, seeks and distinguishes interested the data. The XSLT and the Xpath technology in the data localization and the transformation aspect superiority be used, and solves the key question in the information extraction: Compilate extraction rule. And has carried on the optimization to the extraction rule, causes it simpler, to be vigorous and healthy and to be general.

【关键词】 XMLXSL信息抽取
【Key words】 xmlxsldataextraction
  • 【文献出处】 计算机与数字工程 ,Computer & Digital Engineering , 编辑部邮箱 ,2007年06期
  • 【分类号】TP311.10
  • 【被引频次】17
  • 【下载频次】249
节点文献中: 

本文链接的文献网络图示:

本文的引文网络