节点文献
中文语音合成系统中的文本标准化方法
Text Normalization In Chinese Text-To-Speech System
【摘要】 文本标准化是对输入文本进行分析 ,生成其中非汉字符号的拼音、节奏等信息的过程。本文提出了一种层次化的、基于外部规则的标准化方法 ,通过规则匹配识别这些符号 ,并给出各种正确信息。本文首先介绍了分析树的概念 ,其次给出构造规则的步骤 ,利用权值控制规则的匹配顺序 ,最后给出实验结果。实验结果表明 :这种方法具有很好的易维护性和可扩展性 ,开放测试的正确率达到 99 76 %。
【Abstract】 Text normalization is a procedure to generate information, such as pronunciation, rhythm and so on, for special symbols correctly. In this paper, a method based on hierarchical, external rules is presented. By matching rules, we can recognize normal special symbols and generate correct information. This paper introduces the concept of analysis tree firstly, then shows the steps of constructing rules and presents the experiment results. The results show that we can achieve easy-maintainability and easy-expandability, and the correct rate of open test is 99.76%.
【Key words】 computer application; Chinese information processing; text normalization; special symbols; external rules;
- 【文献出处】 中文信息学报 ,Journal of Chinese Information Processing , 编辑部邮箱 ,2003年04期
- 【分类号】TN912.3
- 【被引频次】17
- 【下载频次】221