节点文献

基于三音子模型连续语音声调识别方法

Multi-feature Based Tri-phone Model for Tone Recognition of Chinese Continuous Speech

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 魏瑞莹梁维谦

【Author】 WEI Ruiyinga,LIANG Weiqianb(a.Department of Microelectronic;b.Department of Electronic Engineering,Tsinghua University,Beijing 100084,China)

【机构】 清华大学微纳电子学系清华大学电子工程系

【摘要】 作为汉语语音识别的重要组成部分,声调识别具有关键的作用。提出了一种新的基于前后文相关的模型识别方法用以提高汉语连续语音中的识别率。首先介绍用于声调识别的基因轨迹的提取和处理,然后提出6种特征来描述基因轨迹的变化趋势并给出具体的计算公式,利用这些特征并考虑连续语音中前后音节的相关性对基因轨迹造成的变化而建立细分的声调模型,最后基于这种声调模型采用决策树的分类方法进行声调的识别和测试。

【Abstract】 As an important part of the recognition of Chinese speech,the tone recognition is also a gut issue.In this paper,a new kind of context-dependent tone model is proposed.Firstly,the Fundamental Frequency(F0) extraction is introduced,and then six kinds of features which contain important information about the discrimination between each type of the tone are proposed.Based on these features and the consideration of the syllable correlations in continuous speech,tri-phone tone models are built.Finally,the decision tree is used to recognize the type of the tone.

  • 【分类号】TN912.34
  • 【被引频次】3
  • 【下载频次】70
节点文献中: 

本文链接的文献网络图示:

本文的引文网络