节点文献
关于外周听觉模型中语音信号处理的分帧问题
【机构】 北京大学视觉与听觉信息处理实验室;
【摘要】 在传统的语音识别系统的前端处理中,基于语音信号的短时平稳特性,对输入的语音信号采取分帧处理的方法。由于听觉系统在对语音信号处理过程中,利用并保持了语音信号的连续性和动态特性。因此在进行听觉建模工作中,分帧的处理方法是否同样适用?本文在计算听觉模型的前端处理中,针对不分帧策略和传统分帧策略哪一种更适合后续处理和提高系统的整体识别性能这一问题,进行了探讨。实验结果表明:在低信噪比情况下,不分帧处理的鲁棒性大大优于分帧处理;在高信噪比情况下,分帧处理的识别结果优于不分帧处理的结果。
【Abstract】 This article compares the effect of framing strategy with that of strategy without framing in the front-end of auditory modeling.We try to answer the question of which strategy is more suitable to the following processing,and can contribute to improving the overall performance of speech recognition system.Our results show that if SNR is low,the strategy without framing is far better than the framing strategy,and if SNR is high,the framing strategy has a better performance.
【Key words】 peripheral auditory model; framing; speech recognition;
- 【会议录名称】 1999年中国神经网络与信号处理学术会议论文集
- 【会议名称】1999年中国神经网络与信号处理学术会议
- 【会议时间】1999-12-01
- 【会议地点】中国广东汕头
- 【分类号】TN912.3
- 【主办单位】中国电子学会、中国神经网络委员会