节点文献

基于U-NET3D的机器人歌声分离

Singing voice separation with U-NET3D for the robot

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 王大东胡希颖王晓宇

【Author】 WANG Da-dong;HU Xi-ying;WANG Xiao-yu;College of Computer,Jilin Normal University;

【机构】 吉林师范大学计算机学院

【摘要】 提出了一种基于U-NET3D的机器人歌声分离方法.为了降低计算复杂度,仅在U-NET3D的第1层使用3维卷积神经网络,从输入的多声道音频中学习不同声源距离产生的幅度和相位特征.利用NAO机器人录制了具有多声源的4声道乐声混合音频数据集,录制的音乐和歌声源自iKala数据集.利用最小欧几里德距离对混音信号、伴奏和歌声进行序列匹配后合成6声道声音数据.实验结果表明,本文所提方法在噪声环境下具有良好的分离效果,与U-NET相比能更好地分离出目标歌声.

【Abstract】 A novel method of singing voice separation with U-NET3D for the robot was proposed.In order to reduce the computation,the 3D convolution neural network was used in the first layer to learn the amplitude and phase characteristics caused by the different distance of source from microphones.The NAO robot was used to record a 4-channel mixed audio dataset with multiple sound sources.The recorded music and singing voice were derived from the iKala dataset.The minimum Euclidean distance was used to match the mixing signal,accompaniment and singing sequence to synthesize 6-channel sound data.The experiment results showed that the proposed method has good separation performance in noisy environment,and can better separate target singing voice than U-NET.

【关键词】 NAO机器人歌声分离U-NET3D
【Key words】 NAO robotsinging voice separationU-NET3D
【基金】 吉林省教育厅“十三五”科学技术规划项目(JJKH20180763K)
  • 【文献出处】 吉林师范大学学报(自然科学版) ,Journal of Jilin Normal University(Natural Science Edition) , 编辑部邮箱 ,2021年01期
  • 【分类号】TP242;TN912.3;TP183;J619.1
  • 【被引频次】1
  • 【下载频次】53
节点文献中: 

本文链接的文献网络图示:

本文的引文网络