节点文献

实体关系模板的获取技术

Extraction of Entity Relation Templates from Text Collections

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 陈晓颖胡熠陆汝占

【Author】 CHEN Xiao-ying,HU Yi,LU Ru-zhan(Department of Computer Science and Engineering,Shanghai Jiaotong University,Shanghai 200240)

【机构】 上海交通大学计算机科学与工程系上海交通大学计算机科学与工程系 上海200240上海200240

【摘要】 确定实体间的关系有助于理解文本,提高信息检索的正确率。该文研究中文实体关系模板的获取技术,提出了一种STG的bootstrapping训练方法。该方法采用生物信息学中的序列比对技术计算上下文的语义模板,使用一定的评估机制筛选模板,有效地扩充元组以提高下一轮训练的质量。实验结果表明,STG生成的模板不仅能覆盖大量的元组,而且正确率可达99%。

【Abstract】 Extracting entity relation is benifit to understand the meaning of text,so as to increase correctness of searching.This paper researches on extracting Chinese entity relation templates from text collections,and puts forward a kind of bootstrapping method called STG.This method makes use of sequence matching technique in bioinformatics to generate semantic templates within context of Chinese entities.A new model of evaluation is presented to select better templates while tuples are expanded to obtain high quality in the next iteration of training.Experimental results show that the templates created by STG not only can cover a large number of tuples,but also can reach 99% accuracy.

【基金】 国家“863”计划基金资助项目(2001AA114210-11);国家自然科学基金资助重大项目(60496326)
  • 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2007年21期
  • 【分类号】TP181
  • 【被引频次】13
  • 【下载频次】168
节点文献中: 

本文链接的文献网络图示:

本文的引文网络