节点文献

多语料库作法之中文姓名辨识

Automatic Recoginition of Chinese Full Name Depending onMultiple corpus

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 张俊盛陈舜德郑萦刘显仲柯淑津

【Author】 Zhang Jun-sheng, Chen Shun-de (Dep. of Information of Tsing hua University, Taiwan.Zheng Ying (Institute of Linguistics of Tsinghua University, Taiwan)Liu Xian-Zhong (Software Division of Sheng bao Institute, Taiwan)Ke Shu-jin (Dep. of Electronic Computer of Dong wu University, Taiwan)

【机构】 台湾清华大学资讯系及语言研究所台湾声宝研究所软体研究室台湾东吴大学电算系

【摘要】 专用名词虽然只占中文文章中的词的百分之一到百分之二,但是,如果不对这些专用名词加以处理,将会形成自动分词的错误的大部分。本文首先描述了包括中文姓名辨识的分词方法,然后介绍其实验结果。最后,文章讨论了中文姓名辨识被遗漏和误判的原因,并提出未来的研究方向。

【Abstract】 The words of non-commonly-used nouns are 1 or 2 percent among the words of Chinese articles, but this will lead to form most of the errors in automatic word segmentation if these words of non-commonly-used nouns are not handled. This paper first describes one automatic word segmenting method including the recoginition of Chinese full name. This paper also introduces its test results. Lastly, this paper discusses the reasons of recoginition of Chinese full name having been lost or wrongly segmented, and puts forward the direction of future research.

  • 【文献出处】 中文信息学报 ,Journal of Chinese Information Processing , 编辑部邮箱 ,1992年03期
  • 【被引频次】67
  • 【下载频次】165
节点文献中: 

本文链接的文献网络图示:

本文的引文网络