节点文献

规则与统计相结合的汉语词义消歧模型

A model of Chinese word sense disambiguation based on combining rule and statistics method

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 谢宇张仰森肖建涛

【Author】 XIE Yu,ZHANG Yang-sen,XIAO Jian-tao(Department of Computer and Automation,Beijing Institute of Machinery,Beijing 100085,China)

【机构】 北京机械工业学院计算机及自动化系北京机械工业学院计算机及自动化系 北京100085北京100085

【摘要】 针对已有的词义消歧研究方法的不足,在分析了多种不同结构知识词典的可计算性及其计算复杂度之后,选择北大计算语言所的《现代汉语语法信息词典》和《现代汉语语义词典》,并结合已经标注了词义的人民日报语料作为词义消歧知识源,从中获取汉语词义消歧所需要的统计知识和规则知识,并采用统计与规则相结合的方法构建词义消歧模型,取得了比较满意的词义消其效果。

【Abstract】 In the light of the disadvantage of word sense disambiguation(WSD) research method,the grammatical knowledge-base of contemporary Chinese and dictionary of temporary Chinese semanitcs are chosen from the Institute of Computational Chinese linguistics of Peking University,after a series of analysis of computability and computational complexity of knowledge dictionary with different structures.Combined the People’s Daily corpus,on which has been tagged word sense,as a word sense disambiguation knowledge source,statistical knowledge and rules knowledge are obtained,which are needed by the Chinese word sense disambiguation.An approach of rule and statistics to construct the model of word sense disambiguation is adopted with satisfactory effect.

【基金】 国家973基金项目(2004CB318102);中国博士后科学基金项目(2005038026)
  • 【文献出处】 北京机械工业学院学报 ,Journal of Beijing Institute of Machinery , 编辑部邮箱 ,2007年03期
  • 【分类号】TP391.1
  • 【被引频次】5
  • 【下载频次】238
节点文献中: 

本文链接的文献网络图示:

本文的引文网络