节点文献

中文金融新闻中公司名的识别

Company Name Identification in Chinese Financial Domain

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王宁葛瑞芳苑春法黄锦辉李文捷

【Author】 WANG Ning 1 GE Rui fang 1 YUAN Chun fa 1 K.F.Wong 2 LI Wen jie 3 (1.State Key Laboratory of Intelligent Technology and System Dept.of Computer Science & Technology Tsinghua University Beijing 100084 2.Dept.of System Engineering & Enginee

【机构】 清华大学计算机科学与技术系香港中文大学系统工程与工程管理系香港理工大学电子计算学系 智能技术与系统国家重点实验室北京100084智能技术与系统国家重点实验室?

【摘要】 在金融领域信息抽取中 ,公司名扮演着非常重要的角色 ;因此如何正确识别文本中出现的公司名是一个非常重要的研究课题。在对金融新闻文本进行了深入地分析和研究的基础上 ,总结出了公司名的结构特征及其上下文信息 ,建立了六个用于识别公司名的知识库 ,并提出了一个基于两次扫描过程的识别策略。初步实验结果表明 ,在封闭测试中实验系统公司名识别的精确率可以达到 97 3% ,召回率可达 89 3% ;在开放测试中精确率可以达到 6 2 8% ,召回率可达 6 2 1%。

【Abstract】 Identifying company names in running texts plays a significant role in financial information extraction.Based on the thoroughly investigations of financial articles,the relevant structural features and contextual constraints were obtained.In this paper,a company name identification system is proposed,which is built on the six knowledge bases and a twice scan method.The experiment achieved 97 3% precision and 89 3% recall respectively by close test,and 62 8% precision and 62 1% recall respectively by open test.

【基金】 国家自然科学基金(6 9975 0 0 8);国家重点基础研究 973(G19980 30 5 0 7)项目支持
  • 【文献出处】 中文信息学报 ,Journal of Chinese Information Processing , 编辑部邮箱 ,2002年02期
  • 【分类号】TP391.4
  • 【被引频次】259
  • 【下载频次】783
节点文献中: 

本文链接的文献网络图示:

本文的引文网络