节点文献

基于MFCC与GFCC混合特征参数的说话人识别

Speaker Recognition Based on Combination of MFCC and GFCC Feature Parameters

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 周萍沈昊郑凯鹏

【Author】 ZHOU Ping;SHEN Hao;ZHENG Kai-peng;College of Electric Engineering and Automation, Guilin University of Electronic Technology;

【机构】 桂林电子科技大学电子工程与自动化学院

【摘要】 针对说话人识别中单一参数表征不够全面的特点,将抗噪性能一般的传统MFCC参数与鲁棒性更强的GFCC参数相互融合,并结合它们的动态特性构成一种新的混合参数.针对特征参数维数过高造成的冗余,研究了每种特征参数各分量与识别结果的关系,舍弃其中贡献较低的分量以实现特征参数降维的目的,并将混合参数应用于基于高斯混合模型的说话人识别系统.仿真实验表明,该混合特征参数具有更好的识别性能和抗噪性.

【Abstract】 Aiming at the issue that single feature parameter of speaker recognition has the shortcoming of low representation ability, a set of mixture feature parameters is formed by combining the single poor anti-noise Mel frequency cepstral coefficients(MFCC) with more robust Gammatone frequency cepstral coefficients(GFCC) and their dynamic differential in this paper. Since the high dimension of the mixture feature parameters, the relationships of each dimension of different feature parameters and recognition results is studied, where dimensionality reduction on high dimensional features is implemented by discarding the dimensions with low contribution ratio. After that, the combination of feature parameters was applied to the speaker recognition system based on Gaussian mixture model. Experimental results show that the combination of parameters can better describe the speakers’ feature and have better anti-noise capability.

【基金】 国家自然科学基金(No.61462017);广西自然科学基金(No.2014GXNSFAA118353);广西自动检测技术与仪器重点实验室基金(No.YQ15110)资助
  • 【文献出处】 应用科学学报 ,Journal of Applied Sciences , 编辑部邮箱 ,2019年01期
  • 【分类号】TN912.34;TN713
  • 【被引频次】49
  • 【下载频次】905
节点文献中: 

本文链接的文献网络图示:

本文的引文网络