节点文献

针对无切分维吾尔文文本行识别的字符模型优化

Character model optimization for segmentation-free Uyghur text line recognition

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 姜志威丁晓青彭良瑞

【Author】 JIANG Zhiwei;DING Xiaoqing;PENG Liangrui;State Key Laboratory of Intelligent Technology and Systems,Tsinghua National Laboratory for Information Science and Technology,Department of Electronic Engineering,Tsinghua University;

【机构】 清华大学电子工程系智能技术与系统国家重点实验室清华信息科学与技术国家实验室

【摘要】 基于隐含Markov模型(hidden Markov model,HMM)的无切分文本行识别方法能够利用概率图的思想,同步完成文本行图像的切分与识别,避免因字符预切分失败而导致的识别错误,但对字符模型的设计与训练要求很高,并且在多字体融合问题中难以提高模型泛化性能。该文通过分析模型状态在图像层面的聚类意义,先提出基于观测合理聚类的模型结构优化方法,再提出结构与参数相结合的字符模型优化策略,最后将其应用于多字体维吾尔文文本行的无切分识别系统。实验结果表明,该方法能够改善模型的状态分配合理性,并且在多字体融合问题中提高了模型泛化性能和状态利用效率。

【Abstract】 A text line recognition method was developed without presegmentation using a hidden Markov model(HMM)for simultaneously segmenting and recognizing text line images.The algorithm uses a probability graph to reduce recognition error from failed presegmentation results.However,the HMM design and training is complicated and the HMM generalization performance can not be easily improved in multi-font texts.Therefore,a character model optimization method with reasonably clustered observations was developed based on the most common HMM state in images.Then,a method was developed to optimize the model structure and parameters together for a multi-font Uyghur text line recognition system.Tests show that this method improves the state allocation,the generalization performance and the state efficiency of the character model for multi-font texts.

【基金】 国家“九七三”重点基础研究项目(2013CB329403)
  • 【文献出处】 清华大学学报(自然科学版) ,Journal of Tsinghua University(Science and Technology) , 编辑部邮箱 ,2015年08期
  • 【分类号】TP391.1
  • 【被引频次】6
  • 【下载频次】119
节点文献中: 

本文链接的文献网络图示:

本文的引文网络