节点文献

隐马尔可夫模型实现语音和视频识别

Hidden Markov Model Based the Recognition of Audio and Video

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 林文永石志国薛为民陈锋军

【Author】 Lin Wen-yong, Shi Zhi-guo, Xue Wei-min, Chen Feng-jun(Information Engineering School, University of Science & Technology Beijing, Beijing 100083 China)

【机构】 北京科技大学信息工程学院

【摘要】 利用隐马尔可夫模型(HMM)对多媒体数据仓库进行复杂数据挖掘,复杂数据挖掘要解决的难题就是音频和视频识别。建立音频和视频的识别模型及其相关的算法,在视频识别算法上,构造出符合HMM的识别方法。根据模型建立系统,实验证明声音的识别率最高达到96.67%,视频中特征值的检测率可以达到87.81%。研究结果可以应用在多媒体的识别和数据挖掘领域,提供一个比较完整的复杂数据挖掘的模型和算法。

【Abstract】 In this paper, it mainly discusses a method to process complex data mining in the multimedia dataware by HMM (Hidden Markov Model). The problem of complex data mining is the recognition of the audio and video. This paper will build the model of recognition of audio and video. It will construct the method according to the HMM. Due to results of experiments, the highest ratio of recognition of audio is 96.67%, the highest ratio of recognition of video is 87.81%. The application area of the results covers recognition of multimedia and the fields of data mining. This paper will provide completely models and technique of CDM (Complex Data Mining).

【基金】 “十五”国家科技攻关项目(2001BA605A)
  • 【会议录名称】 第一届学生计算语言学研讨会论文集
  • 【会议名称】第一届学生计算语言学研讨会
  • 【会议时间】2002-08
  • 【分类号】TN912.3
  • 【主办单位】中国中文信息学会
节点文献中: