节点文献

基于状态异步DBN的语音驱动面部动画合成

Speech Driven Facial Animation Synthesis Based on State Asynchronous DBN

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 赵勇蒋冬梅Sahli Hichem

【Author】 ZHAO Yong;JIANG Dong-mei;Sahli Hichem;School of Computer Science, Northwestern Polytechnical University;ETRO Department, Vrije Universiteit Brussel;

【机构】 西北工业大学计算机学院布鲁塞尔自由大学电子与信息工程系

【摘要】 提出一种基于状态异步动态贝叶斯网络模型(SA-DBN)的语音驱动面部动画合成方法。提取音视频语音数据库中音频的感知线性预测特征和面部图像的主动外观模型(AAM)特征来训练模型参数,对于给定的输入语音,基于极大似然估计原理学习得到对应的最优AAM特征序列,并由此合成面部图像序列和面部动画。对合成面部动画的主观评测结果表明,与听视觉状态同步的DBN模型相比,通过限制听觉语音状态和视觉语音状态间的最大异步程度,SA-DBN可以得到清晰自然并且嘴部运动与输入语音高度一致的面部动画。

【Abstract】 An audio visual Dynamic Bayesian Network model with State Asynchrony(SA-DBN) transforming acoustic speech to photo realistic facial animation is proposed. Perceptual Linear Prediction(PLP) features from audio speech, as well as Active Appearance Model(AAM) features from face images of an audio visual speech database, are adopted to train the model parameters of the proposed SA-DBN. Based on the SADBN model, an input audio stream is given, the optimal AAM visual features are learned by the Maximum Likelihood Estimation(MLE) criterion, which are used to construct facial images for the animation. Subjective evaluation is presented to compare the proposed constrained state asynchrony DBN with a state synchronous audio visual DBN model. Experimental results show that with the SA-DBN model, high quality facial animations can be obtained with mouth movements matching the input speech.

【基金】 国家自然科学基金资助项目(61273265);陕西省国际科技合作基金资助重点项目(2011KW-04)
  • 【文献出处】 计算机工程 ,Computer Engineering , 编辑部邮箱 ,2014年02期
  • 【分类号】TP391.41
  • 【被引频次】2
  • 【下载频次】57
节点文献中: 

本文链接的文献网络图示:

本文的引文网络