节点文献

企业相关信息抽取技术研究与系统实现

Study of the Extracting of the Corporation Attribute Information and a System Implementation

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 张丙奇姜吉发

【Author】 ZHANG Bing qi JIANG Ji fa(Institute of Computing Technology,The Chinese Academy of Sciences,Beijing100080China)

【机构】 中国科学院计算技术研究所软件室中国科学院计算技术研究所软件室 北京100080北京100080

【摘要】 从企业网页中抽取与企业相关的信息是商业上的实际需求,与之相关的研究既有挑战,又有理论意义。文章提出了一个能对中文网页中企业的各种不同类型的属性信息进行抽取的模型,并实现了一个企业相关属性信息抽取系统—CAIES。对该系统进行的测试结果统计表明,它不仅能够满足从网上获取企业竞争情报的实际需求,而且具有较高的抽取正确率与精确率。

【Abstract】 To extract the corporation attribute information from the Web pages of different corporation websites is a factual business demand and the researching about it is also a challenge to us.This paper discusses the key techniques used in the process of extracting these different kinds of corporation attribute information and introduced the design and implementation of an information extraction system-CAIES(Corporation Attributes Information Extraction System).Experiments show that CAIES can do well in extracting different kinds of corporation attribute information.

【关键词】 押信息抽取抽取模型
【Key words】 Information ExtractionExtracting Pattern
  • 【文献出处】 微电子学与计算机 ,Microelectronics & Computer , 编辑部邮箱 ,2004年01期
  • 【分类号】TP393.092
  • 【被引频次】23
  • 【下载频次】160
节点文献中: 

本文链接的文献网络图示:

本文的引文网络