节点文献

基于HMM与词典的汉维词对齐研究

Research on Chinese-Uyghur Word Alignment Based on HMM and Lexicon

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 李萍杨勇任鸽赛买提·艾力

【Author】 LI Ping;YANG Yong;SAI Mai Ti·Ai Li;REN Ge;College of Computer Science and Technology,Xinjiang Normal University;

【机构】 新疆师范大学计算机科学技术学院

【摘要】 词对齐被广泛的用于基于短语的统计机器翻译中,词对齐效果的好坏直接影响了机器翻译的质量。提出将隐马尔科夫模型用于汉维词对齐时,由于汉维双语标记的数据量比较大而且标记数据也还没有公开,导致汉维词对齐的质量较差,也没有办法进行评价,提出采用基于词典的方法进行对齐评价,实现汉维双语词典的构建系统,实验表明,该方法的效果较好,并同时构建汉维双语语料库。

【Abstract】 Word alignment is widely used in statistical machine translation phrase based on phrase. The effect of word alignment directly affects the quality of machine translation. Puts forward using a hidden Markov model for Chinese-Uyghur word alignment, because of the large amount of bilingual marker data and the lack of labeled data, resulting in poor quality of Chinese Uyghur word alignment, there is no way to evaluate. Puts forward the evaluation method based on the alignment dictionary and constructs a bilingual dictionary system. The experiment shows that the effect is good and the Chinese Uighur bilingual corpus is constructed.

【关键词】 隐马尔科夫模型词对齐词典语料库
【Key words】 Hidden Markov ModelWord AlignmentLexiconCorpus
【基金】 新疆师范大学优秀青年教师科研启动基金项目(No.XJNU201420)
  • 【文献出处】 现代计算机(专业版) ,Modern Computer , 编辑部邮箱 ,2017年31期
  • 【分类号】TP391.1
  • 【被引频次】2
  • 【下载频次】59
节点文献中: 

本文链接的文献网络图示:

本文的引文网络