节点文献

用线性预测编码合成改善气管食管语言

Improving tracheoesophageal speech by using LPC synthesis

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 齐颖勇张家■

【Author】 QI Yingyong(Department of Speech and Hearing Sciences, University of Arizona)ZHANG Jialu(Institute of Acoustics, Academia Sinica)

【机构】 亚利桑那大学言语和听觉科学系中国科学院声学研究所 北京 100080

【摘要】 利用气管食管发音是喉切除病人的主要发音方法。本文研究利用线性预测(LPC)语言合成技术,来替换气管食管发音的声源,取得了较好的结果。利用LPC自相关算法分析了四位(二男二女)气管食管发音人所发的四个主要元音[i],[a],[ε],[u]。用归一化预测误差函数来选择分析控制参数和算法。从归一化误差接近最小的那些帧来计算声道传递函数的极点,这些极点是由正常发音导出的。利用重建的声道传递函数和合成的激励声源来合成这些元音。清晰度试验结果表明,合成元音的清晰度很高,并且合成元音的性别分辨要比原来的气管食管发音好得多。

【Abstract】 The feasibility of using the linear predictive coding (LPC) technique to replace the voicing sources of tracheoesophageal speech was explored. Four vowels, [i] , [a] , [] , [u] , and one diphthong [ou] , produced by two male and two female tracheoesophageal speakers were analyzed by the LPC autocorrelation method. Normalized prediction error functions were used to choose the algorithm and the control parameters of the analysis. Poles of the vocal tract transfer function were selected from frames whose normalized prediction errors were close to minimum with criteria derived from transfer functions of normally produced vowels. Vowels were synthesized with the reconstructed transfer function and a synthesized excitation input. Results of an identification task indicated that the synthesized vowels were highly intelligible and the gender of the speaker was better identified from the synthesized vowels than from the original tracheoesophageal vowels.

  • 【被引频次】4
  • 【下载频次】19
节点文献中: 

本文链接的文献网络图示:

本文的引文网络