节点文献

基于聚合经验模态分解的情感语音特征提取

Feature Extraction of Emotional Speech Based on Ensemble Empirical Mode Decomposition

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 张乐张雪英孙颖张卫

【Author】 ZHANG Le;ZHANG Xueying;SUN Ying;ZHANG Wei;College of Information Engineering,Taiyuan University of Technology;

【机构】 太原理工大学信息工程学院

【摘要】 特征提取是情感语音识别系统的关键过程,决定系统整体识别性能。传统特征提取技术假定语音信号是线性、短时平稳信号,不具有自适应性。为此,通过聚合经验模态分解(EEMD)算法以非线性的处理方式提取特征。情感语音信号经EEMD分解后得到一组固有模态函数(IMF),利用相关系数法筛选出有效分量集合,对集合函数计算得到IMF能量特征(IMFE)。选用德国柏林语音库作为实验数据来源,将IMFE特征、韵律特征、梅尔倒谱系数特征以及三者的融合特征分别输入到支持向量机中,通过比较不同特征的识别结果验证IM FE特征的有效性。实验结果表明,IM FE特征与声学特征融合后的平均识别率达到91.67%,可有效区分不同的情感状态。

【Abstract】 Extracting features of emotional speech signal is particularly important in the emotional speech recognition systems,which determines the overall recognition performance. The traditional feature extraction techniques assume speech signal is linear and short-stationary,without self-adapability. By using the Ensemble Empirical Mode Decomposition( EEMD) algorithm,the features are extracted in a nonlinear way. First,the emotional speech signal is decomposed into a series of Intrinsic Mode Function( IMF) by EEMD and effective IMFs set is selected using correlation coefficient method. Then the IMF Energy( IMFE) characteristics are obtained through calculation of the function in the set. In the experiment,Berlin speech database is chosen as the data source. IMFE features,prosodic features,MelFregurecy Cepstrum Coefficients( MFCC) features and the fusion features of the three are input inte SVMrespectively.The recognition results of different feature combinations are compared to validate the performance of the IMFE features.The experimental results showthat the average recognition rate of IMFE feature merging with acoustic feature can reach91. 67%,and IMFE can effectively distingwish between different states.

【基金】 国家自然科学基金(61371193);山西省回国留学人员科研基金(2013-034)
  • 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2017年08期
  • 【分类号】TN912.34
  • 【被引频次】6
  • 【下载频次】180
节点文献中: 

本文链接的文献网络图示:

本文的引文网络