节点文献

稀疏线性预测字典在语音压缩感知中的应用

The application of sparse linear prediction dictionary to compressive sensing in speech signals

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 游寒旭李为李昕朱杰

【Author】 YOU Hanxu;LI Wei;LI Xin;ZHU Jie;School of Electronic Information and Electrical Engineering,Shanghai Jiao Tong University;

【机构】 上海交通大学电子信息与电气工程学院

【摘要】 压缩感知理论框架可以同时实现信号的采样和压缩,将压缩感知应用于语音信号处理是近年来的研究热点之一.本文根据语音信号的特点,采用K-SVD算法获得稀疏线性预测字典,作为语音信号的稀疏变换矩阵.高斯随机矩阵用于原语音信号的采样从而实现信号的压缩,最后通过正交匹配追踪算法(OMP)和采样压缩匹配追踪算法(Co Sa MP)将已采样压缩的语音信号进行信号重构.实验考察了待处理语音信号帧的长度、压缩比,稀疏变换字典以及压缩感知重构算法等因素对语音压缩感知重构性能的影响,结果表明,基于数据集训练的稀疏线性预测字典相比传统解析构造的离散余弦变换字典,对语音的重构性能具有0.6 d B左右的提升.

【Abstract】 Appling compressive sensing( CS),which theoretically guarantees that signal sampling and signal compression can be achieved simultaneously,into audio and speech signal processing is one of the most popular research topics in recent years. In this paper,K-SVD algorithm was employed to learn a sparse linear prediction dictionary regarding as the sparse basis of underlying speech signals. Compressed signals was obtained by applying random Gaussian matrix to sample original speech frames. Orthogonal matching pursuit( OMP) and compressive sampling matching pursuit( Co Sa MP) were adopted to recovery original signals from compressed one. Numbers of experiments were carried out to investigate the impact of speech frames length,compression ratios,sparse basis and reconstruction algorithms on CS performance. Results show that sparse linear prediction dictionary can advance the performance of speech signals reconstruction compared with discrete cosine transform( DCT) matrix.

【基金】 国家自然科学基金(61271349,61371147,11433002);上海交通大学医工合作基金(YG2012ZD04)
  • 【文献出处】 上海师范大学学报(自然科学版) ,Journal of Shanghai Normal University(Natural Sciences) , 编辑部邮箱 ,2016年02期
  • 【分类号】TN912.3
  • 【被引频次】3
  • 【下载频次】110
节点文献中: 

本文链接的文献网络图示:

本文的引文网络