节点文献

一种视频文本自动定位、跟踪和识别的方法

An Algorithm of Automatic Video Text Locating, Tracking and Recognition

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 李朝晖余英林

【Author】 LI Zhao-hui~(1), 2)), YU Ying-lin~(2)) ()~(1))(Department of computer, Information School, Guangzhou University, Guangzhou 510405) ()~(2))(Department of Communication and Electronic Engineering, South China University of Technology, Guangzhou 510641)

【机构】 广州大学信息学院计算机系华南理工大学电子与通信工程系 广州510405华南理工大学电子与通信工程系广州510641广州510641

【摘要】 视频数据中的文本能提供重要的语义信息。本文提出了一种视频文本自动定位、跟踪和识别的方法,首先用基于小波和LH检测视频帧文本所在的位置,然后用运动估计的方法,跟踪后继帧文本的位置,再用多帧平均的方法增强文本区域,最后经过二值化处理和连通分量分析,将文本字符送入OCR软件进行识别。实验结果表明,该方法简单易行,能快速地定位和跟踪文本区域,定位精度和识别效果良好。

【Abstract】 Text in video can provide an important supplemental source of index semantic information. In this paper, an algorithm of automatic video text locating, tracking and recognition is presented. First, the text regions are located by several steps: wavelet decomposition, high frequency component intensity and density detection, horizontal and vertical convex detection based LH, and text locating. Then the text regions are tracked in next consecutive frames. After multiple frames averaging, the text regions are enhanced. By binarization of the enhanced text regions followed by component analysis, the text regions with clean background are obtained. Then the text regions are recognized by OCR software, the final text strings are attained. Experimental results show that the proposed algorithm can detect and track text region simply and effectively.

【关键词】 文本检测语义内容视频索引
【Key words】 text detectingsemantic contextvideo index
【基金】 国家自然科学基金项目(60372068);广东省科学基金项目(011628)
  • 【文献出处】 中国图象图形学报 ,Journal of Image and Graphics , 编辑部邮箱 ,2005年04期
  • 【分类号】TP391.41
  • 【被引频次】22
  • 【下载频次】353
节点文献中: 

本文链接的文献网络图示:

本文的引文网络