节点文献

汉语耳语音孤立字识别研究

Isolated word recognition in Chinese whispered speech

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 杨莉莉林玮徐柏龄

【Author】 YANG Li-Li LIN Wei XU Bo-Ling (Key Laboratory of Modern Acoustics, The Institute of Acoustics, Nanjing University, Nanjing 210093)

【机构】 南京大学声学所近代声学重点实验室南京大学声学所近代声学重点实验室 南京 210093南京 210093

【摘要】 耳语音识别有着广泛的应用前景,是一个全新的课题。但是由于耳语音本身的特点,如声级低、没有基频等,给耳语音识别研究带来了困难。本文根据耳语音信号发音模型,结合耳语音的声学特性,建立了一个汉语耳语音孤立字识别系统。由于耳语音信噪比低,必须对其进行语音增强处理, 同时在识别系统中应用声调信息提高了识别性能。实验结果说明了MFCC结合幅值包络可作为汉语耳语音自动识别的特征参数,在小字库内用HMM模型识别得出的识别率为90.4%。

【Abstract】 The whispered speech recognition is a new subject which has wide applications. However, the characteristics of whispered speech such as its low sound pressure level and the lack of fundamental frequency bring difficulty to the whispered speech recognition. In this paper, a Chinese isolated word recognition system is established based on the source-filter generation model combined with the acoustic characteristics of whispered speech. In addition, the speech enhancement algorithm is added to the system to improve the SNR of whispered speech, and the tone information is implemented to acquire better recognition performance. The experimental results demonstrate that the MFCC combined with the amplitude contour features can be used as efficient parameters for the Chinese whispered speech recognition. A recognition rate of 90.4% is obtained when a small Chinese isolated word database is tested using HMM approach.

【关键词】 耳语音语音识别语音增强
【Key words】 Whispered speechSpeech recognitionSpeech enhancement
【基金】 国家自然科学基金项目(60272037和60340420325)
  • 【分类号】H11
  • 【被引频次】18
  • 【下载频次】184
节点文献中: 

本文链接的文献网络图示:

本文的引文网络