节点文献

一种基于权重矩阵的临近词检索问题解决框架

Weigh Matrix Based Solution Framework for Term Proximity Information Retrieval

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 乔亚男齐勇史椸侯迪王晓

【Author】 QIAO Ya-nan1 QI Yong1 SHI Yi1 HOU Di1 WANG Xiao2(Dept.of Computer Science & Engineering,Xi’an Jiaotong University,Xi’an 710049,China)1(Tangdu Hospital,the Fourth Military Medical University,Xi’an 710038,China)2

【机构】 西安交通大学电信学院计算机系第四军医大学唐都医院

【摘要】 传统的信息检索模型假设查询中的关键词之间是并列关系,但用户的需求往往应该被抽象为一系列的关键词组,组内的关键词间具有更为紧密的语义关系,这就是定义的临近词检索问题。提出了基于权重矩阵的临近词检索问题解决框架,该框架将文档和查询抽象化为文档的权重矩阵表示和查询权重矩阵,通过计算两个矩阵间的相似度来实现临近词检索。实验结果证明,针对临近词检索问题,传统的信息检索模型只是一种简化问题的解决方案,权重矩阵框架从理论上和形式上更加契合临近词检索问题,查准率得到了显著的提高。

【Abstract】 Tradional information retrieval models assume that keywords in queries are parallel,but the requirements of users should be abstracted to a series of keywords groups,and the sematic relations of keywords inside the group are closer than outside.This is "Term Proximity Information Retrieval"(TPIR) defined in this paper,and we presented a solution framework based on Weigh Matrix(WMSF).WMSF abstractes documents and queries to Weigh Matrix Representation of Document and Query Weigh Matrix,and then implements the TPIR based on the caculating of similarity between them.Empirical results show that WMSF is appropriate for TPIR compared with traditional information retrieval models which simplify the TPIR problems actually

【基金】 863基金项目(2006AA01Z101);教育部博士点基金(20060698018);陕西省科技攻关项目(2006K04-G23)资助
  • 【文献出处】 计算机科学 ,Computer Science , 编辑部邮箱 ,2009年07期
  • 【分类号】TP391.3
  • 【被引频次】3
  • 【下载频次】117
节点文献中: 

本文链接的文献网络图示:

本文的引文网络