节点文献

相关向量机及在说话人识别应用中的研究

Study to Speaker Recognition Using RVM

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 杨成福章毅

【Author】 YANG Cheng-fu1,2 and ZHANG Yi1(1.Computational Intelligence Laboratory,University of Electronic Science and Technology of China Chengdu 610054;2.Officer of Education Administrator,Sichuan University of Arts and Science Dazhou Sichuan 635000)

【机构】 电子科技大学计算智能实验室四川文理学院理工系

【摘要】 对基于相关向量机和高斯混合模型的说话人识别算法的模型和特征空间进行了一系列的研究。与一些基于语音帧的说话人识别算法相比,该算法将GMM算法作为底层的语音特征提取,从而实现对语音整体上的处理,对常用的两种语音特征美尔频率倒频系数和瞬时频率的表现进行了对比研究;同时,该算法充分利用了相关向量机的所提供的高泛化性、核函数功能和结果的高稀疏性。基于Chains和AHUMADA两个专门用于说话人识别的语音库的仿真表明,该算法在减少相对误差和减少计算量方面有较大的优势。

【Abstract】 A series of studies on speaker recognition algorithm based on relevance vector machine(RVM) and gaussian mixture model(GMM) was proposed in this paper.The sparseness and probability prediction of RVM make the algorithm suitable for speaker recognition in applications.The robust speech features based on GMM are investigated.In contrast to the most current systems based on frame-level discrimination,the approach has two outstanding merits.The first is the system provides direct discrimination between whole sequences by combining GMM as underlying generative models in feature-space.The paper focused on two main feature space:mel-frequency cepstrum coefficient(MFCC) and instantaneous frequencies(IF).The second combines the high generalization,kernel tricks,and sparser performance of RVM to generate more robust classification results and to reduce the computational complexity.The simulations using the Chains database and the AHUMADA database show that the proposed algorithm outperforms the other systems on reducing the relative error rates and reducing the computational complexity in high dimensionality space and big scale data.

【基金】 国家863计划(2007AA01Z321);四川省教育厅自然科学重点项目(08ZA037)
  • 【文献出处】 电子科技大学学报 ,Journal of University of Electronic Science and Technology of China , 编辑部邮箱 ,2010年02期
  • 【分类号】TN912.34
  • 【被引频次】27
  • 【下载频次】308
节点文献中: 

本文链接的文献网络图示:

本文的引文网络